Activations as Features: Probing LLMs for Generalizable Essay Scoring Representations
Jinwei Chi, Ke Wang, Yu Chen, Xuanye Lin, Qiang Xu
摘要
Automated essay scoring (AES) is a challenging task in cross-prompt settings due to the diversity of scoring criteria. While previous studies have focused on the output of large language models (LLMs) to improve scoring accuracy, we believe activations from intermediate layers may also provide valuable information. To explore this possibility, we evaluated the discriminative power of LLMs’ activations in cross-prompt essay scoring task. Specifically, we used activation to fit probes and further analyzed the effects of different models and input content of LLMs on this discriminative power. By computing the directions of essays across various trait dimensions under different prompts, we analyzed the variation in evaluation perspectives of large language models concerning essay types and traits. Results show that the activations possess strong discriminative power in evaluating essay quality and that LLMs can adapt their evaluation perspectives to different traits and essay types, effectively handling the diversity of scoring criteria in cross-prompt settings.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper8
- Inference-Time Intervention: Eliciting Truthful Answers from a Language ModelKenneth Li, Oam Patel, Fernanda B. Viégas, Hanspeter Pfister 等NeurIPS 2023 · 被引用 1,549 次
- Refusal in Language Models Is Mediated by a Single DirectionAndy Arditi, Oscar Obeso, Aaquib Syed, Daniel Paleka 等NeurIPS 2024 · 被引用 1,166 次
- Automated Cross-prompt Scoring of Essay TraitsRobert Ridley, Liang He, Xin-Yu Dai, Shujian Huang 等AAAI 2021 · 被引用 101 次
- Adaptive Activation Steering: A Tuning-Free LLM Truthfulness Improvement Method for Diverse Hallucinations CategoriesTianlong Wang, Xianfeng Jiao, Yinghao Zhu, Zhongzhi Chen 等WWW 2025 · 被引用 64 次
- PMAES: Prompt-mapping Contrastive Learning for Cross-prompt Automated Essay ScoringYuan Chen, Xia LiACL 2023 · 被引用 20 次
相关 Paper
- KAES: Multi-aspect Shared Knowledge Finding and Aligning for Cross-prompt Automated Scoring of Essay TraitsXia Li, Wenjing PanAAAI 2025 · 被引用 5 次
- LCES: Zero-shot Automated Essay Scoring via Pairwise Comparisons Using Large Language ModelsTakumi Shibata, Yuichi MiyamuraEMNLP 2025 · 被引用 3 次
- Conundrums in Cross-Prompt Automated Essay Scoring: Making Sense of the State of the ArtShengjie Li, Vincent NgACL 2024 · 被引用 8 次
- Cross-Prompt Automated Essay Scoring of Multiple Traits: Making Sense of the State of the ArtShengjie Li, Vincent NgACL 2026 · 被引用 10 次
- Improving Domain Generalization for Prompt-Aware Essay Scoring via Disentangled Representation LearningZhiwei Jiang, Tianyi Gao, Yafeng Yin, Meng Liu 等ACL 2023 · 被引用 16 次
