The Unreasonable Effectiveness of Reasoning Distillation: using DeepSeek R1 to beat OpenAI o1
Source
youtube.com
Author
Latent Space
Date
Why it matters
Shows that distilling reasoning traces from an open reasoning modelA model trained to think — generating extended internal reasoning before answering — trading time and tokens for accuracy on hard problems.Full definition → can lift smaller models substantially, and how data curation drives the result.
Terms in this piece · Glossary
reasoning model — A model trained to think — generating extended internal reasoning before answering — trading time and tokens for accuracy on hard problems.