
If you run judges or collect preference data on cited answers, citation count and diversity are quietly moving the scores — and model judges weigh them differently from humans, without seeing the sources.
articleSelf- and Other-Labels Induce Bidirectional Bias in LLM JudgesSongeun Chae, Min Kim, Donghoon Jung, Seojin Choi, Seohyon Jung
articleWho Do Language Models Think Is Competent? A Mechanistic Analysis of Occupational BiasKeren Fuentes, Aaron Mueller
articleAI Evaluation Should Work With HumansJan Kulveit, Gavin Leech, Tom\'a\v{s} Gaven\v{c}iak, Raymond DouglasChecking sign-in…
Loading comments…