CEC-Zero: Zero-Supervision Character Error Correction with Self-Generated Rewards
Zhiming Lin, Kai Zhao, Sophie Zhang, Peilai Yu, Canran Xiao
Abstract
Large-scale Chinese spelling correction (CSC) remains critical for real-world text processing, yet existing LLMs and supervised methods lack robustness to novel errors and rely on costly annotations. We introduce CEC-Zero, a zerosupervision reinforcement learning framework that addresses this by enabling LLMs to correct their own mistakes. CEC-Zero synthesizes errorful inputs from clean text, computes cluster-consensus rewards via semantic similarity and candidate agreement, and optimizes the policy with PPO. It outperforms supervised baselines by 10-13 F1 points and strong LLM fine-tunes by 5-8 points across 9 benchmarks, with theoretical guarantees of unbiased rewards and convergence. CEC-Zero establishes a label-free paradigm for robust, scalable CSC, unlocking LLM potential in noisy text pipelines.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 967ed013-041d-41b8-b470-8fc356fc192bCited by top-tier papers8
- EEO-TFV: Escape-Explore Optimizer for Web-Scale Time-Series Forecasting and Vision AnalysisHua Wang, Jinghao Lu, Fan ZhangWWW 2026 · 6 citations
- One Refiner to Unlock Them All: Inference-Time Reasoning Elicitation via Reinforcement Query RefinementYixiao Zhou, Dongzhou Cheng, Zhiliang Wu, Yi Yang et al.ACL 2026 · 3 citations
- Synergy over Discrepancy: A Partition-Based Approach to Multi-Domain LLM Fine-TuningHua Ye, Siyuan Chen, Haoliang Zhang, Weihao Luo et al.NeurIPS 2025 · 2 citations
- Whose Instructions Count? Resolving Preference Bias in Instruction Fine-TuningJiayu Zhang, Changbang Li, Yinan Peng, Weihao Luo et al.NeurIPS 2025 · 2 citations
- Aligning by Misaligning: Boundary-aware Curriculum Learning for Multimodal AlignmentHua Ye, Hang Ding, Siyuan Chen, Yiyang Jiang et al.NeurIPS 2025
Builds on18
- TTRL: Test-Time Reinforcement LearningYuxin Zuo, Kaiyan Zhang, Li Sheng, Shang Qu et al.NeurIPS 2025 · 249 citations
- Spelling Error Correction with Soft-Masked BERTShaohua Zhang, Haoran Huang, Jicong Liu, Hang LiACL 2020 · 204 citations
- NDC-Scene: Boost Monocular 3D Semantic Scene Completion in Normalized Device Coordinates SpaceJiawei Yao, Chuming Li, Keqiang Sun, Yingjie Cai et al.ICCV 2023 · 150 citations
- RL on Incorrect Synthetic Data Scales the Efficiency of LLM Math Reasoning by Eight-FoldAmrith Setlur, Saurabh Garg, Xinyang Geng, Naman Garg et al.NeurIPS 2024 · 143 citations
- Swift Sampler: Efficient Learning of Sampler by 10 ParametersJiawei Yao, Chuming Li, Canran XiaoNeurIPS 2024 · 52 citations
Related papers
- A Simple yet Effective Training-free Prompt-free Approach to Chinese Spelling Correction Based on Large Language ModelsHouquan Zhou, Zhenghua Li, Bo Zhang, Chen Li et al.EMNLP 2024 · 2 citations
- Chinese Spelling Correction as Rephrasing Language ModelLinfeng Liu, Hongqiu Wu, Hai ZhaoAAAI 2024 · 36 citations
- A Training-free LLM-based Approach to General Chinese Character Error CorrectionHouquan Zhou, Bo Zhang, Zhenghua Li, Ming Yan et al.ACL 2025
- ARM: An Alignment-and-Replacement Module for Chinese Spelling Check Based on LLMsChangchun Liu, Kai Zhang, Junzhe Jiang, Zirui Liu et al.EMNLP 2024 · 3 citations
- C-LLM: Learn to Check Chinese Spelling Errors Character by CharacterKunting Li, Yong Hu, Liang He, Fandong Meng et al.EMNLP 2024 · 9 citations
