Inception
inception/mercury-2.5
Published by the vendor; the platform bills at the target's declared price.
Inception inception/mercury-2.5
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
Cached input reads $0.004 per 1M
No routing target resolves to this model yet. Declare one and the router picks it up without retraining.
# targets.yaml - declare a target that runs Mercury 2.5
targets:
- id: mercury-2-5
kind: llm
name: Mercury 2.5
metadata:
model: inception/mercury-2.5 # links the target to this catalogue entry
capabilities:
domains: [general]
supports_tools: true
context_window: 260000
cost:
usd_per_1k_tokens: 0.000095Input list price per million tokens; Mercury 2.5 is the cheapest of 2 listed.
Other models by the same vendor in the catalogue.