Unsupervised Pretraining for Fact Verification by Language Model Distillation
Adrián Bazaga, Pietro Lio, Gos Micklem
摘要
Fact verification aims to verify a claim using evidence from a trustworthy knowledge base. To address this challenge, algorithms must produce features for every claim that are both semantically meaningful, and compact enough to find a semantic alignment with the source information. In contrast to previous work, which tackled the alignment problem by learning over annotated corpora of claims and their corresponding labels, we propose SFAVEL (Self-supervised F act V erification via Language Model Distillation), a novel unsupervised pretraining framework that leverages pre-trained language models to distil self-supervised features into highquality claim-fact alignments without the need for annotations. This is enabled by a novel contrastive loss function that encourages features to attain high-quality claim and evidence alignments whilst preserving the semantic relationships across the corpora. Notably, we present results that achieve a new state-of-the-art on FB15k-237 (+5.3% Hits@1) and FEVER (+8% accuracy) with linear evaluation.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- ClaimDB: A Fact Verification Benchmark over Large Structured DataMichael Theologitis, Preetam Prabhu Srikar Dammu, Chirag Shah, Dan SuciuACL 2026 · 被引用 2 次
- IRIS: An Iterative and Integrated Framework for Verifiable Causal Discovery in the Absence of Tabular DataTao Feng, Lizhen Qu, Niket Tandon, Gholamreza HaffariACL 2025
它引用的顶会 Paper20
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- A Simple Framework for Contrastive Learning of Visual RepresentationsTing Chen, Simon Kornblith, Mohammad Norouzi, Geoffrey E. HintonICML 2020 · 被引用 24,064 次
- Unsupervised Learning of Visual Features by Contrasting Cluster AssignmentsMathilde Caron, Ishan Misra, Julien Mairal, Priya Goyal 等NeurIPS 2020 · 被引用 5,249 次
- Big Self-Supervised Models are Strong Semi-Supervised LearnersTing Chen, Simon Kornblith, Kevin Swersky, Mohammad Norouzi 等NeurIPS 2020 · 被引用 2,611 次
- Contrastive Representation DistillationYonglong Tian, Dilip Krishnan, Phillip IsolaICLR 2020 · 被引用 1,305 次
相关 Paper
- Reasoning Over Semantic-Level Graph for Fact CheckingWanjun Zhong, Jingjing Xu, Duyu Tang, Zenan Xu 等ACL 2020 · 被引用 154 次
- Fine-grained Fact Verification with Kernel Graph Attention NetworkZhenghao Liu, Chenyan Xiong, Maosong Sun, Zhiyuan LiuACL 2020 · 被引用 9 次
- CFEVER: A Chinese Fact Extraction and VERification DatasetYing-Jia Lin, Chun-Yi Lin, Chia-Jen Yeh, Yi-Ting Li 等AAAI 2024 · 被引用 8 次
- CLEVE: Contrastive Pre-training for Event ExtractionZiqi Wang, Xiaozhi Wang, Xu Han, Yankai Lin 等ACL 2021
- Counterfactual Debiasing for Fact VerificationWeizhi Xu, Qiang Liu, Shu Wu, Liang WangACL 2023 · 被引用 26 次
