← model directory
inclusionai
Ling-3.0-flash
inclusionai/ling-3.0-flash
⚙ tool calling
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
specs & pricing
- Type
- text
- Provider
- inclusionai
- Model ID
- inclusionai/ling-3.0-flash
- Capabilities
- tools, reasoning
- Context window
- 262K tokens
- Input price
- $0.023 / 1M tokens
- Output price
- $0.069 / 1M tokens
Cloud price is billed from prepaid credits when a request fails over to the cloud. Open-weight models run free on GPUs you own — the gateway routes to your nodes first.