Model catalogue

inclusionAI
Ling 3.0 Flash
inclusionai/ling-3.0-flash
Specification
Published by the vendor; the platform bills at the target's declared price.
Ling 3.0 Flash
reasoning open weights toolsinclusionAI inclusionai/ling-3.0-flash
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
- Input $/1M
- $0.02
- Output $/1M
- $0.06
- Context
- 262K
- Max output
- 33K
- Released
- Jul 23, 2026
- Modalities
- text
Intelligence index27
Coding index51
Agentic index21
Cached input reads $0.004 per 1M · weights on Hugging Face as inclusionAI/Ling-3.0-flash
On this deployment
No routing target resolves to this model yet. Declare one and the router picks it up without retraining.
Not routable here yet
Add a target with metadata.model set to this id.
targets.yamlyaml
# targets.yaml - declare a target that runs Ling 3.0 Flash
targets:
- id: ling-3-0-flash
kind: llm
name: Ling 3.0 Flash
metadata:
model: inclusionai/ling-3.0-flash # links the target to this catalogue entry
capabilities:
domains: [general]
supports_tools: true
context_window: 262144
cost:
usd_per_1k_tokens: 0.000042Price among inclusionAI models
Input list price per million tokens; Ling 3.0 Flash is the 3rd cheapest of 4 listed.
- 1free
- 2free
- 3$0.02Ling 3.0 Flash
- 4$0.06
More from inclusionAI
Other models by the same vendor in the catalogue.