Skip to content
OpenSmartRoute
Model catalogue
inclusionAI

inclusionAI

Ling 3.0 Flash

inclusionai/ling-3.0-flash

Specification

Published by the vendor; the platform bills at the target's declared price.

inclusionAI

Ling 3.0 Flash

reasoning open weights tools

inclusionAI inclusionai/ling-3.0-flash

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

Input $/1M
$0.02
Output $/1M
$0.06
Context
262K
Max output
33K
Released
Jul 23, 2026
Modalities
text
Intelligence index27
Coding index51
Agentic index21

Cached input reads $0.004 per 1M · weights on Hugging Face as inclusionAI/Ling-3.0-flash

On this deployment

No routing target resolves to this model yet. Declare one and the router picks it up without retraining.

Not routable here yet
Add a target with metadata.model set to this id.
targets.yamlyaml
# targets.yaml - declare a target that runs Ling 3.0 Flash
targets:
  - id: ling-3-0-flash
    kind: llm
    name: Ling 3.0 Flash
    metadata:
      model: inclusionai/ling-3.0-flash          # links the target to this catalogue entry
    capabilities:
      domains: [general]
      supports_tools: true
      context_window: 262144
    cost:
      usd_per_1k_tokens: 0.000042

Price among inclusionAI models

Input list price per million tokens; Ling 3.0 Flash is the 3rd cheapest of 4 listed.