
A deep technical breakdown of how DeepSeek achieved frontier-level models at dramatically lower training cost, plus the GPU supply/export-control realities shaping who can build large clusters — useful for anyone reasoning about model economics and compute strategy.
“OpenAI had such little compute and they devoted all of their compute for many months, all of it, 100% for many months to GPT-4 with a brand-new architecture”
Dylan Patel
“The big winners throughout human history are the ones who are willing to do YOLO at some point.”
Lex Fridman
“And language models crash the cost of very intelligent sounding language.”
Dylan Patel
“DeepSeek is doing fantastic work for disseminating understanding of AI. Their papers are extremely detailed in what they do and for other teams around the world, they're very actionable in terms of improving your own training techniques.”
Nathan Lambert
“They freaking have F-35s and we don't let them buy GPUs.”
Dylan Patel
Checking sign-in…
Loading comments…