The new GPT-5.6 Sol powers all chats for paid users, including Instant, creating
Source
OpenAI
Author
OpenAITop Viber
Published
Terms in this piece · Glossary
eval — A repeatable test for AI quality — a set of tasks plus scoring — used the way software teams use test suites, because model output is too variable to judge by eyeballing.
Why it matters
A quantified factuality improvement on the model backing all paid chat traffic is the kind of number builders use when deciding what to route to.
Transcript
The new GPT-5.6 Sol powers all chats for paid users, including Instant, creating one consistent experience.
In our high-stakes factuality evaluation covering finance, medicine and law, the new GPT‑5.6 Sol produced 68% fewer responses with factual errors than GPT‑5.5 Instant. https://t.co/8iKN49tsxc