context window — The maximum amount of text a model can consider at once — its working memory for the current conversation or task.
token — The chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.
Why it matters
Clarifies that 'Qwen 3.8' is four differently licensed models (two downloadable, two hosted-only) with different context windowThe maximum amount of text a model can consider at once — its working memory for the current conversation or task.Full definition → windows and per-provider prices, directly affecting which one you can self-host versus must call via API.