# Ling 3.0 Tiny is now available on AI Gateway

Source: [Vercel](https://vercel.com/changelog/ling-3-0-tiny-is-now-available-on-ai-gateway)  
Feed7 permalink: https://feed7.dev/p/ling-3-0-tiny-is-now-available-on-ai-gateway-1v4j94m  
Published: 2026-08-06T00:00:00.000Z  
Trust: Official Source (official_source)

## Why Included

Ling 3.0 Tiny gives coding agents a small MoE option with native function calling, prompt caching, a 256K context window, and gateway-based routing controls.

## Source Summary

**Ling 3.0 Tiny** is a mixture-of-experts model with **7.9B total parameters**, about **1.3B active per token**, a 256K-token context window, and up to 32K output tokens. It includes native function calling and prompt caching.

## Practical Implication

Builders can select inclusionai/ling-3.0-tiny-free through the AI SDK or Vercel's coding-agent setup. Account for the scheduled switch to inclusionai/ling-3.0-tiny on August 14 when configuring persistent agent environments.

## Agent-Ready Context

**Ling 3.0 Tiny** is a mixture-of-experts model with **7.9B total parameters**, about **1.3B active per token**, a 256K-token context window, and up to 32K output tokens. It includes native function calling and prompt caching.

Builders can select inclusionai/ling-3.0-tiny-free through the AI SDK or Vercel's coding-agent setup. Account for the scheduled switch to inclusionai/ling-3.0-tiny on August 14 when configuring persistent agent environments.

The free model slot ends at 8:00am PT on August 14 and replaces Ling 3.0 Flash. The material describes intended responsiveness but provides no coding evals or production reliability results.

## Connected Context

Feed7 judgment across 390 accumulated Signals:

This replaces Ling 3.0 Flash’s temporary evaluation slot with a smaller mixture-of-experts route that keeps 256K context while adding native function calling, prompt caching, and a scheduled model-ID transition. It narrows adoption to planned testing and configuration migration: parameter counts and intended responsiveness do not establish coding quality or production reliability, consistent with prior candidates’ case for workload-level evaluation.

- [Ling 3.0 Flash is now available on AI Gateway](https://feed7.dev/p/ling-3-0-flash-is-now-available-on-ai-gateway-1he7mve) — Tiny directly replaces Flash in the free slot, preserving the 256K-context evaluation opportunity while requiring persistent environments to change model IDs on the stated date.
- [Laguna S 2.1 is now available on AI Gateway](https://feed7.dev/p/laguna-s-2-1-is-now-available-on-ai-gateway-1nkdv05) — Laguna provides a directly comparable 256K free coding route plus a paid 1M option, so Tiny’s context size alone does not distinguish it for selection.
- [DeepSeek V4 Flash now runs updated weights on AI Gateway](https://feed7.dev/p/deepseek-v4-flash-now-runs-updated-weights-on-ai-gateway-1qqe8mw) — DeepSeek’s behavior changed behind an unchanged ID, while Ling requires an explicit scheduled ID switch; together they show that persistent routing needs both configuration tracking and repeated workload evaluation.
- [AI Gateway adds unified fast mode support](https://feed7.dev/p/ai-gateway-adds-unified-fast-mode-support-144dq26) — Gateway fast mode offers a separate latency control, so Ling’s intended responsiveness should not be treated as measured speed or as a substitute for verifying the serving tier and workload results.

## Context Map

- Layer: model
- Domains: coding
- Topics: open-models, coding-agents, model-selection

## Uncertainty

- The free model slot ends at 8:00am PT on August 14 and replaces Ling 3.0 Flash. The material describes intended responsiveness but provides no coding evals or production reliability results.

## Agent Instruction

Use this item as source-backed context. Do not invent claims beyond the linked source. If this item conflicts with another source, call out the conflict.
