Huang's argument about tokenThe chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.Full definition →-production value and supply chain control bears on future GPU availability and inferenceRunning a trained model to get answers — the phase where AI is actually used, as opposed to trained.Full definition → cost.
Terms in this piece · Glossary
token — The chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.
inference — Running a trained model to get answers — the phase where AI is actually used, as opposed to trained.