Vibeleaderboard
← All Intel
Intel / video

⚡️Mercury: Ultra-Fast Diffusion LLMs — Estefano Ermon, CEO Inception Labs

Source
youtube.com
Author
Latent Space
Date
Why it matters

Diffusion LLMs generate many in parallel, which raises throughput well above autoregressive models. If coding quality holds, loops get faster and cheaper.

Terms in this piece · Glossary
  • token — The chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.
  • AI agent — An AI system that doesn't just answer once but works toward a goal in a loop — taking actions, reading the results, and deciding what to do next.
Read the source www.youtube.com
More from Latent Space
Recommended reads
Comments

Checking sign-in…

Loading comments…