← All IntelIntel / article
Real Time Speech To Text Under 200ms
- Source
- elevenlabs.io
- Author
- ElevenLabs editorial sitemap
- Date
Why it matters
Gives concrete latency levers for voice pipelines, such as about 100 ms PCM chunks, TCP head-of-line blocking versus UDP media, and when to override VAD with manual commits.
Terms in this piece · Glossary
- streaming — Sending a model's response token by token as it is generated, so the reader sees text immediately instead of waiting for the whole answer.
Read the source elevenlabs.io
More from ElevenLabs editorial sitemap
Recommended reads
Comments
Checking sign-in…
Loading comments…