Hi - I answer from the OpenSmartRoute documentation: routing, the API, plans and quotas, self-hosting. Ask away, or open a support ticket if you need a person.
Grounded in the docs - follow a source before acting on it.
Mercury 2.5 - Model - OpenSmartRoute
Modelv1.0.0
Mercury 2.5
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, a
Reference model card from the OpenRouter listing - model inception/mercury-2.5. Prices are per million tokens as listed; ratings and installs below are from this marketplace.
Mercury 2.5
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
Facts
Vendor
Inception
Context window
260,000 tokens
Max output
65,536 tokens
Input price
$0.04 / 1M tokens
Output price
$0.15 / 1M tokens
Input
text
Output
text
Tool calling
yes
Structured outputs
yes
Reasoning
yes
Open weights
no
Endpoint: https://openrouter.ai/api/v1 (OpenAI-compatible), model inception/mercury-2.5
Use it
Copy one of these into your project. Installing also returns the manifest and these snippets.
# after Install: the listing is in your workspace's routing pool - nothing else to configure
curl -s -X POST https://api.opensmartroute.ai/api/v1/route -H 'Authorization: Bearer $OSR_API_KEY' -H 'Content-Type: application/json' -d '{"text": "...", "plan": true}'
# or pin it on the OpenAI-compatible endpoint: {"model": "model-inception-mercury-2-5", ...}
Manifest
An Open Capability Manifest: the router reads it to know what this does, what it costs and when to pick it.