Generating Handwritten Mathematical Expressions From Symbol Graphs: An End-to-End Pipeline
Yu Chen, Fei Gao, Yanguang Zhang, Maoying Qiao, Nannan Wang
摘要
In this paper, we explore a novel challenging generation task, i.e. Handwritten Mathematical Expression Generation (HMEG) from symbolic sequences. Since symbolic sequences are naturally graph-structured data, we formulate HMEG as a graph-to-image (G2I) generation problem. Unlike the generation of natural images, HMEG requires critic layout clarity for synthesizing correct and recognizable formulas, but has no real masks available to supervise the learning process. To alleviate this challenge, we propose a novel end-to-end G2I generation pipeline (i.e. graph → layout → mask → image), which requires no real masks or nondifferentiable alignment between layouts and masks. Technically, to boost the capacity of predicting detailed relations among adjacent symbols, we propose a Less-is-More (LiM) learning strategy. In addition, we design a differentiable layout refinement module, which maps bounding boxes to pixel-level soft masks, so as to further alleviate ambiguous layout areas. Our whole model, including layout prediction, mask refinement, and image generation, can be jointly optimized in an end-to-end manner. Experimental results show that, our model can generate highquality HME images, and outperforms previous generative methods. Besides, a series of ablations study demonstrate effectiveness of the proposed techniques. Finally, we validate that our generated images promisingly boosts the performance of HME recognition models, through data augmentation. Our code and results are available at: https: //github.com/AiArt-HDU/HMEG .
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper1
问问它们各自怎么用它它引用的顶会 Paper26
- LoRA: Low-Rank Adaptation of Large Language ModelsEdward J. Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu 等ICLR 2022 · 被引用 18,833 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Adding Conditional Control to Text-to-Image Diffusion ModelsLvmin Zhang, Anyi Rao, Maneesh AgrawalaICCV 2023 · 被引用 6,759 次
- Vector Quantized Diffusion Model for Text-to-Image SynthesisShuyang Gu, Dong Chen, Jianmin Bao, Fang Wen 等CVPR 2022 · 被引用 607 次
- Diffusion Autoencoders: Toward a Meaningful and Decodable RepresentationKonpat Preechakul, Nattanat Chatthee, Suttisak Wizadwongsa, Supasorn SuwajanakornCVPR 2022 · 被引用 276 次
相关 Paper
- Graph-to-Graph: Towards Accurate and Interpretable Online Handwritten Mathematical Expression RecognitionJin-Wen Wu, Fei Yin, Yan-Ming Zhang, Xu-Yao Zhang 等AAAI 2021 · 被引用 39 次
- Read Ten Lines at One Glance: Line-Aware Semi-Autoregressive Transformer for Multi-Line Handwritten Mathematical Expression RecognitionWentao Yang, Zhe Li, Dezhi Peng, Lianwen Jin 等ACM MM 2023 · 被引用 6 次
- SSAN: A Symbol Spatial-Aware Network for Handwritten Mathematical Expression RecognitionHaoran Zhang, Xiangdong Su, Xingxiang Zhou, Guanglai GaoAAAI 2025 · 被引用 4 次
- Structure-aware Mathematical Expression Recognition with Sequence-Level ModelingMinli Li, Peilin Zhao, Yifan Zhang, Shuaicheng Niu 等ACM MM 2021 · 被引用 4 次
- Handwritten Mathematical Expression Recognition via Attention Aggregation Based Bi-directional Mutual LearningXiaohang Bian, Bo Qin, Xiaozhe Xin, Jianwu Li 等AAAI 2022 · 被引用 70 次
