Towards A Scanpath-Conditioned Surprisal Theory: Modeling Reader Information States
Michael Mooney, Edmond S. L. Ho
摘要
Standard surprisal is typically computed from the linear text prefix, but human reading is non-linear and memory-constrained: readers skip words, regress, and do not retain prior context perfectly. We propose a formulation of surprisal conditioned on a reader-specific accessible information state given by the scanpath history and memory dynamics, rather than by the written prefix alone. Prior context is treated as only probabilistically accessible at each fixation, allowing predictability to depend on both non-linear exposure and forgetting. We evaluate the approach on eye-tracking corpora using held-out log-likelihood over standard duration-based reading measures. Across model variants, conditioning on accessible information states improves predictive fit over standard surprisal baselines. These results suggest that predictability in human reading is better characterized relative to the reader’s evolving accessible information state than to the written prefix alone.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper7
- Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and InferenceBenjamin Warner, Antoine Chaffin, Benjamin Clavié, Orion Weller 等ACL 2025 · 被引用 552 次
- Masked Language Modeling and the Distributional Hypothesis: Order Word Matters Pre-training for LittleKoustuv Sinha, Robin Jia, Dieuwke Hupkes, Joelle Pineau 等EMNLP 2021 · 被引用 177 次
- Masked Language Model ScoringJulian Salazar, Davis Liang, Toan Q. Nguyen, Katrin KirchhoffACL 2020 · 被引用 167 次
- The Impact of Token Granularity on the Predictive Power of Language Model SurprisalByung-Doh Oh, William SchulerACL 2025 · 被引用 7 次
- Information Value: Measuring Utterance Predictability as Distance from Plausible AlternativesMario Giulianelli, Sarenne Wallbridge, Raquel FernándezEMNLP 2023 · 被引用 4 次
相关 Paper
- A Spatio-Temporal Point Process for Fine-Grained Modeling of Reading BehaviorFrancesco Ignazio Re, Andreas Opedal, Glib Manaiev, Mario Giulianelli 等ACL 2025 · 被引用 2 次
- Probing for Reading TimesEleftheria Tsipidi, Samuel Kiegeland, Francesco Ignazio Re, Tianyang Xu 等ACL 2026
- Déjà Vu? Decoding Repeated Reading from Eye MovementsYoav Meiri, Omer Shubi, Cfir Avraham Hadar, Ariel Kreisberg Nitzav 等ACL 2025 · 被引用 1 次
- Entropy- and Distance-Based Predictors From GPT-2 Attention Patterns Predict Reading Times Over and Above GPT-2 SurprisalByung-Doh Oh, William SchulerEMNLP 2022 · 被引用 13 次
- Suspense in Short Stories is Predicted By Uncertainty Reduction over Neural Story RepresentationDavid Wilmot, Frank KellerACL 2020 · 被引用 2 次
