Counterfactual Evaluation of Peer-Review Assignment Policies
Martin Saveski, Steven Jecmen, Nihar B. Shah, Johan Ugander
摘要
Peer review assignment algorithms aim to match research papers to suitable expert reviewers, working to maximize the quality of the resulting reviews. A key challenge in designing effective assignment policies is evaluating how changes to the assignment algorithm map to changes in review quality. In this work, we leverage recently proposed policies that introduce randomness in peer-review assignment--in order to mitigate fraud--as a valuable opportunity to evaluate counterfactual assignment policies. Specifically, we exploit how such randomized assignments provide a positive probability of observing the reviews of many assignment policies of interest. To address challenges in applying standard off-policy evaluation methods, such as violations of positivity, we introduce novel methods for partial identification based on monotonicity and Lipschitz smoothness assumptions for the mapping between reviewer-paper covariates and outcomes. We apply our methods to peer-review data from two computer science venues: the TPDP'21 workshop (95 papers and 35 reviewers) and the AAAI'22 conference (8,450 papers and 3,145 reviewers). We consider estimates of (i) the effect on review quality when changing weights in the assignment algorithm, e.g., weighting reviewers' bids vs. textual similarity (between the review's past papers and the submission), and (ii) the"cost of randomization", capturing the difference in expected quality between the perturbed and unperturbed optimal match. We find that placing higher weight on text similarity results in higher review quality and that introducing randomization in the reviewer-paper assignment only marginally reduces the review quality. Our methods for partial identification may be of independent interest, while our off-policy approach can likely find use evaluating a broad class of algorithmic matching systems.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper3
- AgentReview: Exploring Peer Review Dynamics with LLM AgentsYiqiao Jin, Qinlin Zhao, Yiyang Wang, Hao Chen 等EMNLP 2024 · 被引用 24 次
- Off-policy Evaluation Beyond Overlap: Sharp Partial Identification Under SmoothnessSamir Khan, Martin Saveski, Johan UganderICML 2024 · 被引用 4 次
- A Principled Approach to Randomized Selection under Uncertainty: Applications to Peer Review and Grant FundingAlexander Goldberg, Giulia Fanti, Nihar B. ShahNeurIPS 2025 · 被引用 3 次
它引用的顶会 Paper5
- Mitigating Manipulation in Peer Review via Randomized Reviewer AssignmentsSteven Jecmen, Hanrui Zhang, Ryan Liu, Nihar B. Shah 等NeurIPS 2020 · 被引用 90 次
- A Novice-Reviewer Experiment to Address Scarcity of Qualified Reviewers in Large ConferencesIvan Stelmakh, Nihar B. Shah, Aarti Singh, Hal Daumé IIIAAAI 2021 · 被引用 37 次
- Off-policy Bandits with Deficient SupportNoveen Sachdeva, Yi Su, Thorsten JoachimsKDD 2020 · 被引用 22 次
- SPECTER: Document-level Representation Learning using Citation-informed TransformersArman Cohan, Sergey Feldman, Iz Beltagy, Doug Downey 等ACL 2020 · 被引用 20 次
- Peer Grading the Peer Reviews: A Dual-Role Approach for Lightening the Scholarly Paper Review ProcessInes Arous, Jie Yang, Mourad Khayati, Philippe Cudré-MaurouxWWW 2021 · 被引用 14 次
相关 Paper
- A One-Size-Fits-All Approach to Improving Randomness in Paper AssignmentYixuan Even Xu, Steven Jecmen, Zimeng Song, Fei FangNeurIPS 2023 · 被引用 11 次
- Making Paper Reviewing Robust to Bid Manipulation AttacksRuihan Wu, Chuan Guo, Felix Wu, Rahul Kidambi 等ICML 2021 · 被引用 27 次
- Vulnerability of Text-Matching in ML/AI Conference Reviewer Assignments to CollusionsJhih-Yi Hsieh, Aditi Raghunathan, Nihar B. ShahUSENIX Security 2025
- A Dataset on Malicious Paper Bidding in Peer ReviewSteven Jecmen, Minji Yoon, Vincent Conitzer, Nihar B. Shah 等WWW 2023 · 被引用 22 次
- Uncovering Latent Biases in Text: Method and Application to Peer ReviewEmaad A. Manzoor, Nihar B. ShahAAAI 2021 · 被引用 42 次
