← All IntelIntel / article
Moka v1
- Source
- million.dev
- Author
- aidenybai
- Date
Why it matters
It's a hard, verifiable demonstration of how far task-specific can go — a 127x smaller student retaining most of the teacher's playing strength in a 101 KB browser-loadable ONNX file — useful reference material for anyone sizing a small model for edge or in-browser .
Terms in this piece · Glossary
- distillation — Training a small, cheap model to imitate a big one's outputs, keeping much of the capability at a fraction of the cost.
- inference — Running a trained model to get answers — the phase where AI is actually used, as opposed to trained.
Read the source million.dev
More from aidenybai
Recommended reads
Comments
Checking sign-in…
Loading comments…


