All Models
Ling 3.0 Flash
Ling-3.0-flash is a 124B-parameter Mixture-of-Experts model with approximately 5.1B parameters active per token. It prioritizes token efficiency and production-scale agentic inference, helping coding and tool-using agents complete more work within constrained latency and serving budgets.
Available Providers (2)
| Provider | Model ID | Input Cost | Output Cost | Context | Max Output | Docs |
|---|---|---|---|---|---|---|
| | inclusionai/ling-3.0-flash-free | $0/MTok | $0/MTok | 256K | 256K | |
| | inclusionai/ling-3.0-flash | $0.06/MTok | $0.18/MTok | 262.1K | 32.8K |
Capabilities
Reasoning
Tool Calling
Attachments
Open Weights
Structured Output