Disentangling Writer and Character Styles for Handwriting Generation
Gang Dai, Yifan Zhang, Qingfeng Wang, Qing Du, Zhuliang Yu, Zhuoman Liu, Shuangping Huang
Abstract
Training machines to synthesize diverse handwritings is an intriguing task. Recently, RNN-based methods have been proposed to generate stylized online Chinese characters. However, these methods mainly focus on capturing a person's overall writing style, neglecting subtle style inconsistencies between characters written by the same person. For example, while a person's handwriting typically exhibits general uniformity (e.g., glyph slant and aspect ratios), there are still small style variations in finer details (e.g., stroke length and curvature) of characters. In light of this, we propose to disentangle the style representations at both writer and character levels from individual handwritings to synthesize realistic stylized online handwritten characters. Specifically, we present the style-disentangled Transformer (SDT), which employs two complementary contrastive objectives to extract the style commonalities of reference samples and capture the detailed style patterns of each sample, respectively. Extensive experiments on various language scripts demonstrate the effectiveness of SDT. Notably, our empirical findings reveal that the two learned style representations provide information at different frequency magnitudes, underscoring the importance of separate style extraction. Our source code is public at: https://github.com/dailenson/SDT.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 13b95e49-f150-4dfd-857c-ccabbc5989cbCited by top-tier papers21
- Expanding Small-Scale Datasets with Guided ImaginationYifan Zhang, Daquan Zhou, Bryan Hooi, Kai Wang et al.NeurIPS 2023 · 84 citations
- Cross-Ray Neural Radiance Fields for Novel-view Synthesis from Unconstrained Image CollectionsYifan Yang, Shuhai Zhang, Zixiong Huang, Yubing Zhang et al.ICCV 2023 · 61 citations
- Beyond Isolated Words: Diffusion Brush for Handwritten Text-Line GenerationGang Dai, Yifan Zhang, Yutao Qin, Qiangya Guo et al.ICCV 2025 · 5 citations
- Order-Level Attention Similarity Across Language Models: A Latent CommonalityJinglin Liang, Jin Zhong, Shuangping Huang, Yunqing Hu et al.NeurIPS 2025 · 4 citations
- DiffInk: Glyph- and Style-Aware Latent Diffusion Transformer for Text to Online Handwriting GenerationWei Pan, Huiguo He, Hiuyi Cheng, Yilin Shi et al.ICLR 2026 · 2 citations
Builds on10
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 2,496 citations
- Contrastive Representation DistillationYonglong Tian, Dilip Krishnan, Phillip IsolaICLR 2020 · 1,305 citations
- Fast Vision Transformers with HiLo AttentionZizheng Pan, Jianfei Cai, Bohan ZhuangNeurIPS 2022 · 321 citations
- Self-Supervised Aggregation of Diverse Experts for Test-Agnostic Long-Tailed RecognitionYifan Zhang, Bryan Hooi, Lanqing Hong, Jiashi FengNeurIPS 2022 · 214 citations
- Unleashing the Power of Contrastive Self-Supervised Visual Models via Contrast-Regularized Fine-TuningYifan Zhang, Bryan Hooi, Dapeng Hu, Jian Liang et al.NeurIPS 2021 · 82 citations
Related papers
- Handwriting TransformersAnkan Kumar Bhunia, Salman H. Khan, Hisham Cholakkal, Rao Muhammad Anwer et al.ICCV 2021 · 64 citations
- Handwritten Text Generation from Visual ArchetypesVittorio Pippi, Silvia Cascianelli, Rita CucchiaraCVPR 2023
- Learning to Generate Stylized Handwritten Text via a Unified Representation of Style, Content, and NoiseHonglie Wang, Yan-Ming Zhang, Wangzi Yao, Fei Yin et al.ICLR 2026
- XMP-Font: Self-Supervised Cross-Modality Pre-training for Few-Shot Font GenerationWei Liu, Fangyue Liu, Fei Ding, Qian He et al.CVPR 2022 · 64 citations
- GAN-Based Unpaired Chinese Character Image Translation via Skeleton Transformation and Stroke RenderingYiming Gao, Jiangqin WuAAAI 2020 · 71 citations
