GPT-6 Astra is our best model yet for computer use, software engineering, and visual understanding to date. Hear from the builders putting it to work:
Enter the AGI Era with GPT-6 Astra. GPT-6 Astra scores 75.2% on DeepSWE v1.1 at xhigh reasoning. Ask Astra to fix a bug and have it: • Reproduce the bug from your report • Trace the cause across files • Compare fixes and side effects • Patch the code causing the failure • Run tests and retry the failing workflow

GPT-6 Astra scores 72.6% on OSWorld 2.0 Offline, which tests desktop tasks without internet access. With computer-use tools, Astra can work in the app itself: navigate menus, enter information, and inspect the result on screen. It can switch to code when the task calls for it.
GPT-6 Astra brings stronger visual judgment to front-end design. Give Astra a sketch, reference, or existing UI, and ask it to: • Turn the reference into a working UI • Refine layout, typography, and spacing • Adjust color and interactions • Use screenshots to guide revisions
A new frontier model with stated coding and desktop-control scores is reaching the OpenAI API and AWS in days, so teams building coding or computer-use agents have a new option to evaluate against their current stack.
Checking sign-in…
Loading comments…