token — The chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.
Why it matters
Alongside chat/completions there is a fill-in-the-middle endpoint for prefix/suffix infilling, and a diffusing flag that streams the intermediate denoising states. Neither has an equivalent in an autoregressive API.