token — The chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.
context window — The maximum amount of text a model can consider at once — its working memory for the current conversation or task.
tool use — A model's ability to call external functions — run code, search the web, edit files — instead of only generating text.
Why it matters
A 1M context windowThe maximum amount of text a model can consider at once — its working memory for the current conversation or task.Full definition → Flash model with thinking on by default, at half price through December 31, reachable from Claude Code, Codex, Cursor and other agents through a single gateway model string.