Sign InOpen Brain
VercelEngineering PostOfficial Source

Qwen 3.8 Flash now available on AI Gateway

Qwen 3.8 Flash is now selectable in Vercel AI Gateway and coding agents, with text-and-image input, a 1M-token context window, and responses up to 65k tokens.

Vercel · Aug 26, 2026
Open Source Open MarkdownOpen JSON
Source Summary

Alibaba’s **Qwen 3.8 Flash** accepts text and images, has a **1 million-token context window**, and can return up to **65k tokens**. Its Gateway model ID is alibaba/qwen3.8-flash.

Practical Implication

Builders can route Claude Code, Codex, Cursor, and other supported agents through AI Gateway, then test this model on coding, tool-use, and multi-step workloads before changing defaults.

Agent-Ready Context
Alibaba’s **Qwen 3.8 Flash** accepts text and images, has a **1 million-token context window**, and can return up to **65k tokens**. Its Gateway model ID is alibaba/qwen3.8-flash.

Builders can route Claude Code, Codex, Cursor, and other supported agents through AI Gateway, then test this model on coding, tool-use, and multi-step workloads before changing defaults.

The workload recommendations come from Alibaba, while the supplied material provides no benchmark, price, latency, or reliability comparison against other coding models.
Connected Context · Feed7 Judgment

Qwen 3.8 Flash expands the million-token multimodal agent pool and adds a concrete 65K output ceiling, making it a plausible candidate for large repository and image-assisted workflows. It does not narrow model selection by itself: provider workload claims and capacity figures still require matched tests against neighboring routes for quality, tool use, latency, price, and reliability.

GLM 5.3 Flash now available on AI GatewayGLM 5.3 Flash matches Qwen’s text-and-vision input and 1M context while explicitly supporting function calls, structured output, and streaming, making it a close interface-level comparison for agent integration.Qwen 3.8 Max now available on Vercel AI GatewayQwen 3.8 Max is the nearest same-family alternative for long-context text-and-vision work, so testing Flash against Max can determine whether the Flash route changes the quality and operating tradeoff within Qwen 3.8.Hy4 Preview now available on AI GatewayHy4 offers another 1M-context route aimed at repository-scale work, but with an open MoE configuration; together they show why context length alone cannot decide the default.DeepSeek V4 Flash Vision Experimental now available on AI GatewayDeepSeek V4 Flash Vision supplies another vision-and-tool-use route but is explicitly experimental, contrasting its stated production risk with Qwen’s announcement while leaving both without comparative performance evidence.
Context Map
modelcodingimage#model-selection#coding-agents#tool-use
Uncertainty
The workload recommendations come from Alibaba, while the supplied material provides no benchmark, price, latency, or reliability comparison against other coding models.