Collects interviews with OpenAI staff on the Realtime API, o1 and distillationTraining a small, cheap model to imitate a big one's outputs, keeping much of the capability at a fraction of the cost.Full definition → announcements, useful context windowThe maximum amount of text a model can consider at once — its working memory for the current conversation or task.Full definition → for engineers building voice and reasoning features on these APIs.
Terms in this piece · Glossary
distillation — Training a small, cheap model to imitate a big one's outputs, keeping much of the capability at a fraction of the cost.
context window — The maximum amount of text a model can consider at once — its working memory for the current conversation or task.