We evaluated GPT-6 Astra on WANDR. It scored 0.682 at $11.98 per task, the highest score of any model we tested. GPT-6-Astra scored 13.5% higher than Fable 5.1 at 6.1% lower cost, and 27.0% higher than Opus 5 at 3.3% higher cost.

GPT-6 Astra leads Perplexity's WANDR at 0.682 for $11.98 per task, 13.5% above Fable 5.1 at 6.1% lower cost and 27.0% above Opus 5 at 3.3% higher cost, giving a direct capability-per-dollar comparison for .
postPerplexity open sources Lily, an Apple silicon engine built for hybrid agents
postPerplexity puts Claude Fable 5.1 top of its WANDR agent evaluationChecking sign-in…
Loading comments…