Learning To Replicate Expert Judgment In Financial Tasks
Source
Thinking Machines editorial sitemap
Author
Thinking Machines editorial sitemap
Date
Terms in this piece · Glossary
distillation — Training a small, cheap model to imitate a big one's outputs, keeping much of the capability at a fraction of the cost.
inference — Running a trained model to get answers — the phase where AI is actually used, as opposed to trained.
Why it matters
A concrete recipe for encoding expert judgment into a small model: high-quality human annotations plus asymmetric-clipping CISPO and on-policy distillationTraining a small, cheap model to imitate a big one's outputs, keeping much of the capability at a fraction of the cost.Full definition →, outperforming frontier models on accuracy and recall at a fraction of inferenceRunning a trained model to get answers — the phase where AI is actually used, as opposed to trained.Full definition → cost.