Sign InOpen Brain
VercelEngineering PostOfficial Source

Ling 3.0 Flash is now available on AI Gateway

Ling 3.0 Flash joins AI Gateway with a 256K context window, thinking and non-thinking modes, and free access through August 3 for agent workload testing.

Vercel · Jul 23, 2026
Open Source Open MarkdownOpen JSON
Source Summary

AI Gateway now offers **Ling 3.0 Flash**, a Mixture-of-Experts model with **124B total and about 5.1B active parameters per token**. It has a **256K-token context window** and is free through **August 3**.

Practical Implication

Builders can use the temporary free endpoint to evaluate high-frequency coding, document, and multi-step agent runs in both thinking and non-thinking modes, ideally against their own latency and token budgets.

Agent-Ready Context
AI Gateway now offers **Ling 3.0 Flash**, a Mixture-of-Experts model with **124B total and about 5.1B active parameters per token**. It has a **256K-token context window** and is free through **August 3**.

Builders can use the temporary free endpoint to evaluate high-frequency coding, document, and multi-step agent runs in both thinking and non-thinking modes, ideally against their own latency and token budgets.

The material provides architecture and positioning but no measured quality, speed, or reliability results. Free availability is time-limited, so avoid treating promotional pricing as the production cost baseline.
Context Map
modelcoding#open-models#model-selection#coding-agents
Uncertainty
The material provides architecture and positioning but no measured quality, speed, or reliability results. Free availability is time-limited, so avoid treating promotional pricing as the production cost baseline.