Cerebras CEO on the Future of Data Centres, Token Costs & Memory | Should US Companies Sell to China
Source
youtube.com
Author
20VC
Date
Why it matters
Persistent memory shortages and a data center backlog constrain how much inferenceRunning a trained model to get answers — the phase where AI is actually used, as opposed to trained.Full definition → capacity exists and what it costs. A chipmaker CEO's view helps engineers plan for tokenThe chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.Full definition → pricing and availability.
Terms in this piece · Glossary
inference — Running a trained model to get answers — the phase where AI is actually used, as opposed to trained.
token — The chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.