⚖️Purpose-built rerankers such as rerank-2.5 >>> LLMs - Up to 60x cheaper & 48x faster than SOTA LLMs - Beat LLM rerankers by up to 15% in accuracy (NDCG@10) - Strong first-stage retrieval + specialized reranker = best reranking quality

❌ Long context windows don't help LLM rerankers. Single-pass reranking significantly underperforms sliding window reranking, undermining the value of longer context windows for reranking.

💰rerank-2.5 costs $0.05 per 1M tokens vs. $1.25-$3 for LLMs such as Gemini 2.5 Pro, GPT-5, and Claude Sonnet 4.5.

⚡rerank-2.5 offers the best reranking performance, while being 9x, 36x, and 48x faster than Claude Sonnet 4.5, GPT-5, and Gemini 2.5 Pro, respectively.

If you rerank with a frontier , this puts numbers on the cost and latency you pay, and reports that a long does not remove the need for sliding window passes.
Checking sign-in…
Loading comments…