← All IntelIntel / article 
LitReview Arena: Evaluating Literature Review Agents with Battle-Style Peer Review Platform
- Source
- arxiv.org
- Author
- Ruotong Zhao, Zhiyu Chen, Xurui Liu, Haidong Xue, Dong Liang, Jigao Fu, Wu YanBiao, Yuanyi Zhen, Fengli Xu, Yong Li
- Date

Why it matters
Puts a number on how far literature-review agents still sit behind expert drafts, and shows the agentic scaffold — not the base model — carries most of the gain.
Terms in this piece · Glossary
- LLM-as-judge — Using one model to score another's output against a rubric, so quality can be measured at a scale human grading cannot reach.
Read the source arxiv.org
Recommended reads
articleXAI-Arena: Can LLMs Assess the Quality of XAI Explanations?Yanfei Hu Fleischhauer, Alona Zharova, Nadja Klein, Stefan Feuerriegel
articleDeepInstructor: An Agentic AI Instructor for Experience-Driven Idea EvaluationRongcan Pei, Fang Guo, Qinglin Qi, Qi Zhu, Yun Luo, Jianhao Yan, Minjun Zhu, Qiujie Xie, Dehong Zheng, Yue Zhang- articleRubricReviewer: From Direct Critique to Objective and Comprehensive Rubric-Driven Peer ReviewShuyu Guo, Wenxiang Hu, Yuyue Zhao, Yougang Lyu, Xiaohui Yan
Comments
Checking sign-in…
Loading comments…