VercelEngineering PostOfficial Source
Ling 3.0 Flash is now available on AI Gateway
Ling 3.0 Flash joins AI Gateway with a 256K context window, thinking and non-thinking modes, and free access through August 3 for agent workload testing.
Vercel · Jul 23, 2026
Source Summary
AI Gateway now offers **Ling 3.0 Flash**, a Mixture-of-Experts model with **124B total and about 5.1B active parameters per token**. It has a **256K-token context window** and is free through **August 3**.
Practical Implication
Builders can use the temporary free endpoint to evaluate high-frequency coding, document, and multi-step agent runs in both thinking and non-thinking modes, ideally against their own latency and token budgets.
Agent-Ready Context
AI Gateway now offers **Ling 3.0 Flash**, a Mixture-of-Experts model with **124B total and about 5.1B active parameters per token**. It has a **256K-token context window** and is free through **August 3**. Builders can use the temporary free endpoint to evaluate high-frequency coding, document, and multi-step agent runs in both thinking and non-thinking modes, ideally against their own latency and token budgets. The material provides architecture and positioning but no measured quality, speed, or reliability results. Free availability is time-limited, so avoid treating promotional pricing as the production cost baseline.
Context Map
modelcoding#open-models#model-selection#coding-agentsUncertainty
The material provides architecture and positioning but no measured quality, speed, or reliability results. Free availability is time-limited, so avoid treating promotional pricing as the production cost baseline.