Smart enough to reason, fast enough for a phone call. On real voice agent prompts, nothing beats Mercury 2 on both. The world's first reasoning diffusion LLM: full reasoning pass in <300ms at 1000+ tok/s on standard NVIDIA GPUs. 💬"The reasoning quality we need without sacrificing the latency for a natural phone call." — Oliver Silverstein, CEO, @OpenCall_AI Read how it works 👉

Building voice agents? Get in touch with our team 👉 https://t.co/i1RLyI9Wlk
Reasoning models have generally been too slow for a live phone call. A sub-300ms reasoning pass at over 1,000 per second on standard GPUs puts a reasoning step inside a single voice turn.
Checking sign-in…
Loading comments…