
The new GPT-5.6 Sol powers all chats for paid users, including Instant, creating one consistent experience. In our high-stakes factuality evaluation covering finance, medicine and law, the new GPT‑5.6 Sol produced 68% fewer responses with factual errors than GPT‑5.5 Instant.

A quantified factuality improvement on the model backing all paid chat traffic is the kind of number builders use when deciding what to route to.
postOpenAI Publishes Disclosure Framework and Six Misalignment Reports
postOpenAI details its AI-driven 'Defense Factory' for vulnerability huntingChecking sign-in…
Loading comments…