
Retrieval for code agents is routinely tuned against recall@k, and this is direct evidence that optimizing that metric can cost you resolved issues when the is fixed.
articleCodeGrep: An RL-Trained Retrieval Agent for LLM Coding AgentsWuya Chen, Yihao yang, Yang Cao, Yue Lin
articleTry Again, Don't Look Back: Blind Resampling Outperforms Self-Repair in Small Code ModelsYuvraj Verma
videoMemory Harnesses for Long-Running Research Agents — Stefania Druga, Sakana.aiAI EngineerSign in to comment.
Loading comments…