← model directory

inclusionai

Ling-3.0-flash

inclusionai/ling-3.0-flash

⚙ tool calling

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

specs & pricing
Type
text
Provider
inclusionai
Model ID
inclusionai/ling-3.0-flash
Capabilities
tools, reasoning
Context window
262K tokens
Input price
$0.023 / 1M tokens
Output price
$0.069 / 1M tokens

Cloud price is billed from prepaid credits when a request fails over to the cloud. Open-weight models run free on GPUs you own — the gateway routes to your nodes first.