All Models

Ling 3.0 Flash

ling Reasoning Tool Calling

Ling-3.0-flash is a 124B-parameter Mixture-of-Experts model with approximately 5.1B parameters active per token. It prioritizes token efficiency and production-scale agentic inference, helping coding and tool-using agents complete more work within constrained latency and serving budgets.

Providers 2
Released Jul 23, 2026
Input Modalities text
Output Modalities text
Tarsk Use coding

Available Providers (2)

Provider Model ID Input Cost Output Cost Context Max Output Docs
Vercel AI Gateway inclusionai/ling-3.0-flash-free $0/MTok $0/MTok 256K 256K
NanoGPT inclusionai/ling-3.0-flash $0.06/MTok $0.18/MTok 262.1K 32.8K

Capabilities

Reasoning
Tool Calling
Attachments
Open Weights
Structured Output