CFEVER: A Chinese Fact Extraction and VERification Dataset
Ying-Jia Lin, Chun-Yi Lin, Chia-Jen Yeh, Yi-Ting Li, Yun-Yu Hu, Chih-Hao Hsu, Mei-Feng Lee, Hung-Yu Kao
Abstract
We present CFEVER, a Chinese dataset designed for Fact Extraction and VERification. CFEVER comprises 30,012 manually created claims based on content in Chinese Wikipedia. Each claim in CFEVER is labeled as "Supports", "Refutes", or "Not Enough Info" to depict its degree of factualness. Similar to the FEVER dataset, claims in the "Supports" and "Refutes" categories are also annotated with corresponding evidence sentences sourced from single or multiple pages in Chinese Wikipedia. Our labeled dataset holds a Fleiss' kappa value of 0.7934 for five-way inter-annotator agreement. In addition, through the experiments with the state-of-the-art approaches developed on the FEVER dataset and a simple baseline for CFEVER, we demonstrate that our dataset is a new rigorous benchmark for factual extraction and verification, which can be further used for developing automated systems to alleviate human fact-checking efforts. CFEVER is available at https://ikmlab.github.io/CFEVER .
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers4
- Truth over Tricks: Measuring and Mitigating Shortcut Learning in Misinformation DetectionHerun Wan, Jiaying Wu, Minnan Luo, Zhi Zeng et al.NeurIPS 2025 · 14 citations
- RaCMC: Residual-Aware Compensation Network with Multi-Granularity Constraints for Fake News DetectionXinquan Yu, Ziqi Sheng, Wei Lu, Xiangyang Luo et al.AAAI 2025 · 9 citations
- How Do Social Bots Participate in Misinformation Spread? A Comprehensive Dataset and AnalysisHerun Wan, Minnan Luo, Zihan Ma, Guang Dai et al.EMNLP 2025 · 3 citations
- TrendFact: A Benchmark Towards Hotspot Perception in Automatic Fact-CheckingXiaocheng Zhang, Xi Wang, Yifei Lu, Jianing Wang et al.ACL 2026
Builds on5
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida et al.NeurIPS 2022 · 24,707 citations
- Deberta: decoding-Enhanced Bert with Disentangled AttentionPengcheng He, Xiaodong Liu, Jianfeng Gao, Weizhu ChenICLR 2021 · 3,729 citations
- Big Bird: Transformers for Longer SequencesManzil Zaheer, Guru Guruganesh, Kumar Avinava Dubey, Joshua Ainslie et al.NeurIPS 2020 · 3,159 citations
- Mining Dual Emotion for Fake News DetectionXueyao Zhang, Juan Cao, Xirong Li, Qiang Sheng et al.WWW 2021 · 332 citations
- Explainable Automated Fact-Checking for Public Health ClaimsNeema Kotonya, Francesca ToniEMNLP 2020 · 10 citations
Related papers
- DeSePtion: Dual Sequence Prediction and Adversarial Examples for Improved Fact-CheckingChristopher Hidey, Tuhin Chakrabarty, Tariq Alhindi, Siddharth Varia et al.ACL 2020 · 7 citations
- DialFact: A Benchmark for Fact-Checking in DialoguePrakhar Gupta, Chien-Sheng Wu, Wenhao Liu, Caiming XiongACL 2022
- Reasoning Over Semantic-Level Graph for Fact CheckingWanjun Zhong, Jingjing Xu, Duyu Tang, Zenan Xu et al.ACL 2020 · 154 citations
- Fine-grained Fact Verification with Kernel Graph Attention NetworkZhenghao Liu, Chenyan Xiong, Maosong Sun, Zhiyuan LiuACL 2020 · 9 citations
- Unsupervised Pretraining for Fact Verification by Language Model DistillationAdrián Bazaga, Pietro Lio, Gos MicklemICLR 2024 · 5 citations
