Powered by @Cerebras, Ultrafast generates up to 750 tokens per second, bringing
Source
OpenAI
Author
OpenAI
Date
Terms in this piece · Glossary
token — The chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.
Why it matters
Frontier-quality output at ~750 tokenThe chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.Full definition →/second changes what fits inside an interactive loop — real-time voice agents and tight coding iterations that currently stall on generation speed.