
The reconsideration and self-refinement loops meant to improve answers instead amplify sycophancy, and more capable models capitulate more — measure this before adding another refinement pass.
articleDiagnostic Foundation for Evaluating LLMs' Research Integrity as Co-ScientistsYash Tripathi, Silu Sharma, Sai Sidhanth Manoharan Jayanthi, Shivank Garg, Lin Li
articleWho Do Language Models Think Is Competent? A Mechanistic Analysis of Occupational BiasKeren Fuentes, Aaron Mueller
articleNot All Tokens Are Equal: Inflation-Aware Routing for Agentic LLM SystemsHeming Fu, Shan Lin, Qianqian Xie, Guojun XiongChecking sign-in…
Loading comments…