Vibeleaderboard
← All Intel
Intel / article

Real Time Speech To Text Under 200ms

Source
elevenlabs.io
Author
ElevenLabs editorial sitemap
Date
Why it matters

Gives concrete latency levers for voice pipelines, such as about 100 ms PCM chunks, TCP head-of-line blocking versus UDP media, and when to override VAD with manual commits.

Terms in this piece · Glossary
  • streaming — Sending a model's response token by token as it is generated, so the reader sees text immediately instead of waiting for the whole answer.
Read the source elevenlabs.io
More from ElevenLabs editorial sitemap
Recommended reads
Comments

Checking sign-in…

Loading comments…