Vendor
inclusionAI models
4 models in the reference catalogue, vendor list prices per million tokens. 0 are routable on this deployment today.
Models
4
0 routable here
Median list price
$0.02
3:1 input to output per 1M tokens
Every inclusionAI model
Newest first. Prices are the vendor's published rates per million tokens; click a model for the full specification.
| Model | Released | Context | Input / 1M | Output / 1M | Intelligence | Capabilities |
|---|---|---|---|---|---|---|
| Ling 3.0 Flash Sante (free)inclusionai/ling-3.0-flash-sante:free | Sep 4, 2026 | 262K | free | free | - | reasoning tools |
| Ling 3.0 Flash Fininclusionai/ling-3.0-flash-fin | Aug 27, 2026 | 262K | $0.06 | $0.18 | - | reasoning tools |
| Ling 3.0 Flash Fin (free)inclusionai/ling-3.0-flash-fin:free | Aug 27, 2026 | 262K | free | free | - | reasoning tools |
| Ling 3.0 Flashinclusionai/ling-3.0-flash | Jul 23, 2026 | 262K | $0.02 | $0.06 | 27 | reasoning tools open weights |
Frequently asked
- What is the cheapest inclusionAI model?
- Ling 3.0 Flash Sante (free) at free per 1M input tokens and free per 1M output tokens (vendor list price).
- Which inclusionAI model has the largest context window?
- Ling 3.0 Flash Sante (free) accepts 262K tokens of context.
- How do I route to inclusionAI models with OpenSmartRoute?
- Declare a target in targets.yaml with metadata.model set to the catalogue id (for example inclusionai/ling-3.0-flash-sante:free); the router scores it against every other target on cost, quality, latency and your policies for each request.
Route inclusionAI with everything else
OpenSmartRoute picks the cheapest model that meets your quality bar per request, so a inclusionAI flagship handles hard prompts while small models take the rest. Price a workload or try a routing decision live.