All Models
Ling 3.0 Flash
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Benchmarks
Available Providers (4)
| Provider | Model ID | Input Cost | Output Cost | Context | Max Output | Docs |
|---|---|---|---|---|---|---|
| | inclusionai/ling-3.0-flash | $0.02/MTok | $0.06/MTok | 262.1K | 32.8K | |
| | inclusionai/ling-3.0-flash | $0.06/MTok | $0.18/MTok | 256K | 32K | |
| | inclusionai/ling-3.0-flash | $0.06/MTok | $0.18/MTok | 262.1K | 32.8K | |
| | inclusionai/ling-3.0-flash | $0.07/MTok | $0.22/MTok | 262.1K | 32.8K |
Capabilities
Reasoning
Tool Calling
Attachments
Open Weights
Structured Output