Qwen 3.8 Max now available on Vercel AI Gateway
Vercel AI Gateway now exposes Qwen 3.8 Max to coding agents, adding one model endpoint for long-context text and vision work with gateway routing, budgets, and usage tracking.
Vercel AI Gateway now serves **Qwen 3.8 Max** under alibaba/qwen3.8-max. The model combines text and vision-language work, has **2.4 trillion parameters**, and supports up to **1 million tokens** of context.
Builders can connect Claude Code, Codex, OpenCode, or Pi through the gateway setup command, then select the model for coding, screenshot-to-page, captioning, or image-grounded tasks. Gateway controls include usage and cost tracking, retries, failover, budgets, routing, and Zero Data Retention support.
Vercel AI Gateway now serves **Qwen 3.8 Max** under alibaba/qwen3.8-max. The model combines text and vision-language work, has **2.4 trillion parameters**, and supports up to **1 million tokens** of context. Builders can connect Claude Code, Codex, OpenCode, or Pi through the gateway setup command, then select the model for coding, screenshot-to-page, captioning, or image-grounded tasks. Gateway controls include usage and cost tracking, retries, failover, budgets, routing, and Zero Data Retention support. The announcement gives specifications and intended use cases, but no quality, latency, or agent benchmark results. The long context and parameter count do not establish whether it is a better default than models already in an agent stack.
This expands the gateway’s coding-model choice with one endpoint spanning text, vision, and very long context. It supports testing consolidated multimodal workflows through existing agent clients, but does not change the prior conclusion that routing should be workload-specific: parameter count and context length offer no comparative evidence on quality, latency, reliability, or cost.