Peer Grading the Peer Reviews: A Dual-Role Approach for Lightening the Scholarly Paper Review Process
Ines Arous, Jie Yang, Mourad Khayati, Philippe Cudré-Mauroux
Abstract
Scientific peer review is pivotal to maintain quality standards for academic publication. The effectiveness of the reviewing process is currently being challenged by the rapid increase of paper submissions in various conferences. Those venues need to recruit a large number of reviewers of different levels of expertise and background. The submitted reviews often do not meet the conformity standards of the conferences. Such a situation poses an ever-bigger burden on the meta-reviewers when trying to reach a final decision. In this work, we propose a human-AI approach that estimates the conformity of reviews to the conference standards. Specifically, we ask peers to grade each other's reviews anonymously with respect to important criteria of review conformity such as sufficient justification and objectivity. We introduce a Bayesian framework that learns the conformity of reviews from both the peer grading process, historical reviews and decisions of a conference, while taking into account grading reliability. Our approach helps meta-reviewers easily identify reviews that require clarification and detect submissions requiring discussions while not inducing additional overhead from reviewers. Through a large-scale crowdsourced study where crowd workers are recruited as graders, we show that the proposed approach outperforms machine learning or review grades alone and that it can be easily integrated into existing peer review systems. CCS CONCEPTS • Information systems → Crowdsourcing; • Mathematics of computing → Bayesian computation; • Computing methodologies → Neural networks; Learning latent representations.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext a7663c0f-57e8-4a8a-9d55-0c4e6e7291acCited by top-tier papers3
- Counterfactual Evaluation of Peer-Review Assignment PoliciesMartin Saveski, Steven Jecmen, Nihar B. Shah, Johan UganderNeurIPS 2023 · 19 citations
- IdeaSynth: Iterative Research Idea Development Through Evolving and Composing Idea Facets with Literature-Grounded FeedbackKevin Pu, K. J. Kevin Feng, Tovi Grossman, Tom Hope et al.CHI 2025 · 15 citations
- HybridEval: A Human-AI Collaborative Approach for Evaluating Design Ideas at ScaleSepideh Mesbah, Ines Arous, Jie Yang, Alessandro BozzonWWW 2023 · 5 citations
Builds on1
Related papers
- A Novice-Reviewer Experiment to Address Scarcity of Qualified Reviewers in Large ConferencesIvan Stelmakh, Nihar B. Shah, Aarti Singh, Hal Daumé IIIAAAI 2021 · 37 citations
- Vulnerability of Text-Matching in ML/AI Conference Reviewer Assignments to CollusionsJhih-Yi Hsieh, Aditi Raghunathan, Nihar B. ShahUSENIX Security 2025
- Can The Crowd Identify Misinformation Objectively?: The Effects of Judgment Scale and Assessor's BackgroundKevin Roitero, Michael Soprano, Shaoyang Fan, Damiano Spina et al.SIGIR 2020 · 2 citations
- CoCoReviewBench: A Completeness- and Correctness-Oriented Benchmark for AI ReviewersHexuan Deng, Xiaopeng Ke, Yichen Li, Ruina Hu et al.ICML 2026
- The AI Review Lottery: Widespread AI-Assisted Peer Reviews Boost Paper Scores and Acceptance RatesGiuseppe Russo, Manoel Horta Ribeiro, Tim R. Davidson, Veniamin Veselovsky et al.CSCW 2025 · 10 citations
