CollageNoter: Real-Time and Adaptive Collage Layout Design for Screenshot-Based E-Note-Taking
Qiuyun Zhang, Bin Guo, Lina Yao, Xiaotian Qiao, Ying Zhang, Zhiwen Yu
摘要
To enhance the processing of complex multi-modal documents (e.g. e-books, long web pages, etc.), it is an efficient way for users to take digital screenshots of key parts and reorganize them into a new collage E-Note. Existing methods for assisting collage layout design primarily employ a semantic relevance-first strategy, with arranging related contents together. Though capable, it can not ensure the visual readability of screenshots and may conflict with human natural reading patterns. In this paper, we introduce CollageNoter for real-time collage layout design that adapts to various devices (e.g. laptop, tablet, phone, etc.), offering users with visually and cognitively well-organized screenshot-based E-Notes. Specifically, we construct a novel two-stage pipeline for collage design, including 1) readability-first layout generation and 2) cognitive-driven layout adjustment. In addition, to achieve real-time response and adaptive model training, we propose a cascade transformer-based layout generator named CollageFormer and a size-aware collage layout builder for automatic dataset construction. Extensive experimental results have confirmed the effectiveness of our CollageNoter.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- LayoutVAE: Stochastic Scene Layout Generation From a Label SetAkash Abdu Jyothi, Thibaut Durand, Jiawei He, Leonid Sigal 等ICCV 2019 · 被引用 194 次
- LayoutTransformer: Layout Generation and Completion with Self-attentionKamal Gupta, Justin Lazarow, Alessandro Achille, Larry Davis 等ICCV 2021 · 被引用 184 次
- GRIDS: Interactive Layout Design with Integer ProgrammingNiraj Ramesh Dayama, Kashyap Todi, Taru Saarelainen, Antti OulasvirtaCHI 2020 · 被引用 65 次
- SoftCollage: A Differentiable Probabilistic Tree Generator for Image CollageJiahao Yu, Li Chen, Mingrui Zhang, Mading LiCVPR 2022 · 被引用 3 次
- Graph Transformer GANs for Graph-Constrained House GenerationHao Tang, Zhenyu Zhang, Humphrey Shi, Bo Li 等CVPR 2023
相关 Paper
- Collaposer: Transforming Photo Collections into Visual Assets for Storytelling with CollagesJiayi Zhou, Liwenhan Xie, Jiaju Ma, Zheng Wei 等CHI 2026 · 被引用 3 次
- Generative Layout Modeling using Constraint GraphsWamiq Para, Paul Guerrero, Tom Kelly, Leonidas J. Guibas 等ICCV 2021 · 被引用 93 次
- M6Doc: A Large-Scale Multi-Format, Multi-Type, Multi-Layout, Multi-Language, Multi-Annotation Category Dataset for Modern Document Layout AnalysisHiuyi Cheng, Peirong Zhang, Sihang Wu, Jiaxin Zhang 等CVPR 2023
- Graphic Design with Large Multimodal ModelYutao Cheng, Zhao Zhang, Maoke Yang, Hui Nie 等AAAI 2025 · 被引用 3 次
- NoiseCollage: A Layout-Aware Text-to-Image Diffusion Model Based on Noise Cropping and MergingTakahiro Shirakawa, Seiichi UchidaCVPR 2024 · 被引用 19 次
