Vibeleaderboard
← All Intel
Intel / video

How Open Source Became AI's Backbone | Inferact with a16z

Source
youtube.com
Author
a16z
Date
Why it matters

vLLM runs on a huge share of GPUs, so its maintainers' view on speed tiers, hardware support and control informs how you choose where to serve models.

Terms in this piece · Glossary
  • inference — Running a trained model to get answers — the phase where AI is actually used, as opposed to trained.
  • open weights — A model whose trained parameters are published for anyone to download and run — unlike API-only models you can access but never possess.
Read the source www.youtube.com
More from a16z
Recommended reads
Comments

Checking sign-in…

Loading comments…