Gracefully Air-Written: Enhancing the Legibility and Style Consistency of In-Air Handwriting
Yu Liu, Cunrui Wang, Lin Feng, Jianxin Zhang, Bo Lu
摘要
Space computing devices expand handwritten input from two-dimensional screens into three-dimensional space, providing an unrestricted interactive experience. Due to the high degree of freedom and lack of tactile feedback in in-air handwriting, handwritten characters not only become less legible but also lose the writer's personal style. This paper proposes a method for reconstructing discrete in-air handwriting using continuous diffusion models, capturing the writing process and style from a small number of user-provided handwritten tracks and images, to restore the legibility of characters and mimics the writer's style. We represent handwritten track data in binary form and model it with continuous diffusion models, recovering discrete handwritten track data through threshold processing. Our approach reconstructs inair handwritten characters in two stages. During the content preservation phase, we propose a partial noise injection strategy based on reference sequence modeling, using the content information of the original character as a guiding condition to maintain content consistency in handwritten character. In the style aggregation phase, we adaptively fuse the visual style of the handwritten in the image modality with the dynamic writing process in the sequence modality, overcoming issues of insufficient style capture due to noise interference in the backward process. Qualitative and quantitative experiments demonstrate the superiority of our method.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper13
- Denoising Diffusion Probabilistic ModelsJonathan Ho, Ajay Jain, Pieter AbbeelNeurIPS 2020 · 被引用 35,902 次
- Scalable Diffusion Models with TransformersWilliam Peebles, Saining XieICCV 2023 · 被引用 5,568 次
- Diffusion-LM Improves Controllable Text GenerationXiang Lisa Li, John Thickstun, Ishaan Gulrajani, Percy Liang 等NeurIPS 2022 · 被引用 1,546 次
- Argmax Flows and Multinomial Diffusion: Learning Categorical DistributionsEmiel Hoogeboom, Didrik Nielsen, Priyank Jaini, Patrick Forré 等NeurIPS 2021 · 被引用 782 次
- DiffuSeq: Sequence to Sequence Text Generation with Diffusion ModelsShansan Gong, Mukai Li, Jiangtao Feng, Zhiyong Wu 等ICLR 2023 · 被引用 94 次
相关 Paper
- AirSketch: Generative Motion to SketchHui Xian Grace Lim, Xuanming Cui, Yogesh S. Rawat, Ser Nam LimNeurIPS 2024 · 被引用 4 次
- Air-Text: Air-Writing and Recognition SystemSun-Kyung Lee, Jong-Hwan KimACM MM 2021 · 被引用 16 次
- Learning to Generate Stylized Handwritten Text via a Unified Representation of Style, Content, and NoiseHonglie Wang, Yan-Ming Zhang, Wangzi Yao, Fei Yin 等ICLR 2026
- InkFlow: Connected Handwriting Recognition for Natural Mid-Air Input in Mixed RealityXufeng Jian, Qi Qi, Linpei Zhang, Haifeng Sun 等CHI 2026 · 被引用 1 次
- DiffInk: Glyph- and Style-Aware Latent Diffusion Transformer for Text to Online Handwriting GenerationWei Pan, Huiguo He, Hiuyi Cheng, Yilin Shi 等ICLR 2026 · 被引用 2 次
