This week we launched Mercury Voice, a dLLM optimized for voice agents with 2x+ lower latency than models like GPT-6 Luna.
In this dental clinic receptionist demo, median end-to-end latency stays under 300ms, tool calls included.
See (or hear?) it yourself in the model playground: https://t.co/NaT8trzyD4
Enterprise customers interested in Mercury Voice can contact us at sales@inceptionlabs.ai to get access.
Mercury Voice is a diffusion LLMA large language model — the neural network behind tools like Claude and ChatGPT, trained on huge amounts of text to predict what comes next.Full definition → claiming 2x+ lower latency than GPT-6 Luna, with median end-to-end under 300ms including tool calls. This matters when latency is the bottleneck in voice agents.
Terms in this piece · Glossary
LLM — A large language model — the neural network behind tools like Claude and ChatGPT, trained on huge amounts of text to predict what comes next.