multimodal — A model that works with more than text — reading images, audio, or video, and sometimes generating them too.
context window — The maximum amount of text a model can consider at once — its working memory for the current conversation or task.
token — The chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.
Why it matters
Gives the pricing, architecture and context windowThe maximum amount of text a model can consider at once — its working memory for the current conversation or task.Full definition →-window numbers for GLM-5.3-Flash against Opus 4.8 and GPT-5.6 Terra, plus a first-hand account of what the METR/Redwood investigation concluded about the OpenAI–Hugging Face incident.