
RAG architecture choice is usually made on vibes; this puts accuracy-versus-cost scaling behind it on comparable footing.
“The results reveal a scale-dependent crossover rather than an unconditional winner.”
“Around 10 million corpus tokens, BM25 overtakes it and leads at every larger shared tier, with a margin approaching 20 points at full scale.”
“Overall, corpus growth increasingly favors global candidate ranking: lexical retrieval is the strongest scalable default, while agentic reasoning works best after ranked discovery rather than in place of it.”
videoThe unreasonable effectiveness of BM25 for agentic search — Jo Kristian Bergum, Hornet.devAI Engineer
articleRetrieval-Augmented Generation for Scientific Code UnderstandingAaron Nobile, Andreas Adelmann, Mohsen Sadr
articleUniversal Pathologies, Conditional Consequences: A Triple-Robustness Analysis of RAG for Multi-Hop TraceabilityMeftun Akarsu, Burak OzdemirChecking sign-in…
Loading comments…