Vibeleaderboard
← All Intel
Intel / post

GPT-6.1 Sol reaches near-Astra intelligence at a quarter of the cost per task

Source
Artificial Analysis
Date
Artificial Analysis@ArtificialAnlys

GPT-6.1 Sol replaces GPT-6 Sol after just 7 days. It scores 1 point below GPT-6 Astra in the Intelligence Index at less than one quarter of the Cost per Task Pricing matches GPT-6 Sol at $2/$10 per million input/output tokens, except that the cache read discount rises from 90% to 95%. GPT-6.1 Sol’s overall blended price for agentic workloads is therefore slightly lower than GPT-6 Sol. This represents an additional price cut, following GPT-6 Sol’s original 50% discount from GPT-5.6 Sol. Key takeaways: ➤ Achieves near-Astra Intelligence: GPT-6.1 Sol gains 4 points in the Intelligence Index vs GPT-6 Sol, and 5 points vs GPT-5.6 Sol - landing 1 point below GPT-6 Astra. It makes significant gains in agentic knowledge work, improving 4 points and 5 points in AA-Briefcase v1.1 and GDPval-AA v2.1 respectively. Other notable gains include a 12 point jump in Terminal-Bench 4.0, a 5 point jump in Humanity’s Last Exam, a 6 point jump in GDP.pdf, and an 8 point jump in AA-Omniscience Accuracy coupled with hallucination rate falling from 60% to 54%. ➤ Pushes cost efficiency frontier: At max effort, GPT-6.1 Sol costs less than a quarter of GPT-6 Astra per Intelligence Index task ($0.72 vs…

Read the full post on X
Key takeaways · AI-distilled
  • Pricing is unchanged from GPT-6 Sol at $2/$10 per million input/output , but the cache read discount rises from 90% to 95%, so Artificial Analysis says blended agentic cost is slightly lower. GPT-6.1 Sol replaces GPT-6 Sol after 7 days.
  • Artificial Analysis measures a 4-point Intelligence Index gain over GPT-6 Sol, with notable jumps of 12 points on Terminal-Bench 4.0, 5 on Humanity's Last Exam, 6 on GDP.pdf and 8 on AA-Omniscience accuracy, as rate fell from 60% to 54%.
  • At max effort it costs $0.72 per Intelligence Index task, 31% less than GPT-6 Sol ($1.05) and 64% less than GPT-5.6 Sol ($1.99). Artificial Analysis says every effort level extends the cost-efficiency Pareto frontier.
  • It uses about 10-30% more output tokens than GPT-6 Sol across effort levels, though its low and medium settings remain Pareto optimal for token efficiency because of the higher scores.
  • On the Artificial Analysis Coding Index, it gains 3 points on GPT-6 Sol at max effort and sits 2 points below GPT-6 Astra.
Terms in this piece · Glossary
  • hallucination — When a model states something false with full confidence — inventing facts, citations, or APIs that don't exist.
  • AI agent — An AI system that doesn't just answer once but works toward a goal in a loop — taking actions, reading the results, and deciding what to do next.
  • token — The chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.
Why it matters

GPT-6.1 Sol lands 1 point below GPT-6 Astra at under a quarter of the cost per task and with a 95% cache read discount. Agent workloads that don't need Astra can run much cheaper.

More from Artificial Analysis
Recommended reads
Comments

Checking sign-in…

Loading comments…