It gives you empirical production numbers — not vendor benchmarks — on which models are actually absorbing token volume versus where the money goes, useful when deciding whether to route cheap open-weight models for bulk work and reserve premium models for the paths that need them.
The per-modality breakdown also shows what real teams are picking for image and video generation.
Vercel's monthly report on AI Gateway traffic finds open-weight models jumped from 11% to 29% of token volume between April and June 2026 while consuming under 4% of spend, with DeepSeek alone reaching 22.6% of tokens, just behind Google.
Anthropic still dominates spend at 61% despite only 32% of token volume, and the report breaks down per-modality leaders (OpenAI's GPT Image, xAI's Grok Imagine) plus a brief note on Claude Fable 5's release and export-control suspension.