token — The chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.
context window — The maximum amount of text a model can consider at once — its working memory for the current conversation or task.
Why it matters
Advertised context windowThe maximum amount of text a model can consider at once — its working memory for the current conversation or task.Full definition → and leaderboard placement both proved unreliable here. Verify long-context behavior on your own hardware before designing around a headline number.