Vibeleaderboard
← All Intel
Intel / post

Inception launches Mercury Voice, a diffusion LLM for low-latency voice agents

Source
x.com
Date
Inception@_inception_ai

This week we launched Mercury Voice, a dLLM optimized for voice agents with 2x+ lower latency than models like GPT-6 Luna. In this dental clinic receptionist demo, median end-to-end latency stays under 300ms, tool calls included. See (or hear?) it yourself in the model playground: https://t.co/NaT8trzyD4 Enterprise customers interested in Mercury Voice can contact us at sales@inceptionlabs.ai to get access.

Why it matters

Mercury Voice is a diffusion claiming 2x+ lower latency than GPT-6 Luna, with median end-to-end under 300ms including tool calls. This matters when latency is the bottleneck in voice agents.

Terms in this piece · Glossary
  • LLM — A large language model — the neural network behind tools like Claude and ChatGPT, trained on huge amounts of text to predict what comes next.
More from Inception
Recommended reads
Comments

Checking sign-in…

Loading comments…