Vibeleaderboard
← All Intel
Intel / article

Qwen 3.8 Max now available on Vercel AI Gateway

Source
vercel.com
Author
Walter Korman
Date
Why it matters

If you're already routing calls through Vercel's AI Gateway, you can now call a very large Qwen model with a huge through the same key, fallback, and tracing setup you use for other providers, without standing up a separate Alibaba Cloud account.

Terms in this piece · Glossary
  • LLM — A large language model — the neural network behind tools like Claude and ChatGPT, trained on huge amounts of text to predict what comes next.
  • token — The chunk of text a model reads and writes in — roughly three-quarters of a word — and the unit AI usage is billed in.
  • context window — The maximum amount of text a model can consider at once — its working memory for the current conversation or task.
  • multimodal — A model that works with more than text — reading images, audio, or video, and sometimes generating them too.
Key quotes

“Qwen 3.8 Max handles text-only and vision-language work in one model, with 2.4 trillion parameters and a context window of up to 1 million tokens.”

“The model is suited for software engineering and office productivity, along with visual work like turning screenshots or design files into working pages, captioning video, and answering questions grounded in an image.”

“AI Gateway reflects provider pricing with no markup and does not charge a platform fee on inference, including on Bring Your Own Key (BYOK) requests.”

More from Walter Korman
Recommended reads
Comments

Checking sign-in…

Loading comments…