BBScoreV2: Learning Time-Evolution and Latent Alignment from Stochastic Representation
Tianhao Zhang, Zhecheng Sheng, Zhexiao Lin, Chen Jiang, Dongyeop Kang
摘要
Autoregressive generative models play a key role in various language tasks, especially for modeling and evaluating long text sequences. While recent methods leverage stochastic representations to better capture sequence dynamics, encoding both temporal and structural dependencies and utilizing such information for evaluation remains challenging. In this work, we observe that fitting transformer-based model embeddings into a stochastic process yields ordered latent representations from originally unordered model outputs. Building on this insight and prior work, we theoretically introduce a novel likelihood-based evaluation metric BB-ScoreV2. Empirically, we demonstrate that the stochastic latent space induces a "clustered-totemporal ordered" mapping of language model representations in high-dimensional space, offering both intuitive and quantitative support for the effectiveness of BBScoreV2. Furthermore, this structure aligns with intrinsic properties of natural language and enhances performance on tasks such as temporal consistency evaluation (e.g., Shuffle tasks) and AIgenerated content detection.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper11
- SimCSE: Simple Contrastive Learning of Sentence EmbeddingsTianyu Gao, Xingcheng Yao, Danqi ChenEMNLP 2021 · 被引用 2,496 次
- DetectGPT: Zero-Shot Machine-Generated Text Detection using Probability CurvatureEric Mitchell, Yoonho Lee, Alexander Khazatsky, Christopher D. Manning 等ICML 2023 · 被引用 988 次
- The emergence of clusters in self-attention dynamicsBorjan Geshkovski, Cyril Letrouit, Yury Polyanskiy, Philippe RigolletNeurIPS 2023 · 被引用 163 次
- On Contrastive Learning for Likelihood-free InferenceConor Durkan, Iain Murray, George PapamakariosICML 2020 · 被引用 149 次
- Language modeling via stochastic processesRose E. Wang, Esin Durmus, Noah D. Goodman, Tatsunori HashimotoICLR 2022 · 被引用 28 次
相关 Paper
- BBScore: A Brownian Bridge Based Metric for Assessing Text CoherenceZhecheng Sheng, Tianhao Zhang, Chen Jiang, Dongyeop KangAAAI 2024 · 被引用 8 次
- Automatic Text Evaluation through the Lens of Wasserstein BarycentersPierre Colombo, Guillaume Staerman, Chloé Clavel, Pablo PiantanidaEMNLP 2021 · 被引用 21 次
- BERTScore: Evaluating Text Generation with BERTTianyi Zhang, Varsha Kishore, Felix Wu, Kilian Q. Weinberger 等ICLR 2020 · 被引用 8,443 次
- Is Everything in Order? A Simple Way to Order SentencesSomnath Basu Roy Chowdhury, Faeze Brahman, Snigdha ChaturvediEMNLP 2021 · 被引用 2 次
- Repeated Sequences Reveal Gaps between Large Language Models and Natural LanguageKumiko Tanaka-IshiiACL 2026 · 被引用 1 次
