A Foundation Model for Error Correction Codes
Yoni Choukroun, Lior Wolf
摘要
In recent years, Artificial Intelligence has undergone a paradigm shift with the rise of foundation models, which are trained on large amounts of data, typically in a self-supervised way, and can then be adapted to a wide range of downstream tasks. In this work, we propose the first foundation model for Error Correction Codes. This model is trained on multiple codes and can then be applied to an unseen code. To enable this, we extend the Transformer architecture in multiple ways: (1) a code-invariant initial embedding, which is also position-and lengthinvariant, (2) a learned modulation of the attention maps that is conditioned on the Tanner graph, and (3) a length-invariant code-aware noise prediction module that is based on the parity-check matrix. The proposed architecture is trained on multiple short-and medium-length codes and is able to generalize to unseen codes. Its performance on these codes matches and even outperforms the state of the art, despite having a smaller capacity than the leading code-specific transformers. The suggested framework therefore demonstrates, for the first time, the benefits of learning a universal decoder rather than a decoder optimized for a given code.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Efficient Message-Passing Transformer for Error Correcting CodesSeong-Joon Park, Taewoo Park, Hee-Youl Kwak, Sang-Hyo Kim 等ICLR 2026 · 被引用 29 次
- Learning Linear Block Error Correction CodesYoni Choukroun, Lior WolfICML 2024 · 被引用 18 次
- Drop-in Circulant Structural Priors for Transformer Decoding of Cyclic CodesShuai Xiao, Weijun Fang, Qiaosheng ZhangICML 2026
- CrossMPT: Cross-attention Message-passing Transformer for Error Correcting CodesSeong-Joon Park, Heeyoul Kwak, Sang-Hyo Kim, Yongjune Kim 等ICLR 2025
- Score Based Error Correcting Code DecoderAlon Helvits, Eliya NachmaniICML 2026
它引用的顶会 Paper9
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- Chain-of-Thought Prompting Elicits Reasoning in Large Language ModelsJason Wei, Xuezhi Wang, Dale Schuurmans, Maarten Bosma 等NeurIPS 2022 · 被引用 22,562 次
- High-Resolution Image Synthesis with Latent Diffusion ModelsRobin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser 等CVPR 2022 · 被引用 13,123 次
- Photorealistic Text-to-Image Diffusion Models with Deep Language UnderstandingChitwan Saharia, William Chan, Saurabh Saxena, Lala Li 等NeurIPS 2022 · 被引用 8,965 次
相关 Paper
- Error Correction Code TransformerYoni Choukroun, Lior WolfNeurIPS 2022 · 被引用 121 次
- Self-Taught Recognizer: Toward Unsupervised Adaptation for Speech Foundation ModelsYuchen Hu, Chen Chen, Chao-Han Huck Yang, Chengwei Qin 等NeurIPS 2024 · 被引用 14 次
- Are Transformers universal approximators of sequence-to-sequence functions?Chulhee Yun, Srinadh Bhojanapalli, Ankit Singh Rawat, Sashank J. Reddi 等ICLR 2020 · 被引用 481 次
- CBraMod: A Criss-Cross Brain Foundation Model for EEG DecodingJiquan Wang, Sha Zhao, Zhiling Luo, Yangxuan Zhou 等ICLR 2025
- Principled Understanding of Generalization for Generative Transformer Models in Arithmetic Reasoning TasksXingcheng Xu, Zibo Zhao, Haipeng Zhang, Yanqing YangACL 2025 · 被引用 2 次
