How Open Source Became AI's Backbone | Inferact with a16z
Source
youtube.com
Author
a16z
Date
Why it matters
vLLM runs on a huge share of inferenceRunning a trained model to get answers — the phase where AI is actually used, as opposed to trained.Full definition → GPUs, so its maintainers' view on speed tiers, hardware support and open weightsA model whose trained parameters are published for anyone to download and run — unlike API-only models you can access but never possess.Full definition → control informs how you choose where to serve models.
Terms in this piece · Glossary
inference — Running a trained model to get answers — the phase where AI is actually used, as opposed to trained.
open weights — A model whose trained parameters are published for anyone to download and run — unlike API-only models you can access but never possess.