Hi - I answer from the OpenSmartRoute documentation: routing, the API, plans and quotas, self-hosting. Ask away, or open a support ticket if you need a person.
Grounded in the docs - follow a source before acting on it.
Nemotron 3 Ultra (free) - Model - OpenSmartRoute
Modelv1.0.0
Nemotron 3 Ultra (free)
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts ar
Reference model card from the OpenRouter listing - model nvidia/nemotron-3-ultra-550b-a55b:free. Prices are per million tokens as listed; ratings and installs below are from this marketplace.
Nemotron 3 Ultra (free)
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Facts
Vendor
NVIDIA
Context window
1,000,000 tokens
Max output
65,536 tokens
Input price
$0.00 / 1M tokens
Output price
$0.00 / 1M tokens
Input
text
Output
text
Tool calling
yes
Structured outputs
no
Reasoning
yes
Open weights
nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16
Endpoint: https://openrouter.ai/api/v1 (OpenAI-compatible), model nvidia/nemotron-3-ultra-550b-a55b:free
Use it
Copy one of these into your project. Installing also returns the manifest and these snippets.
# after Install: the listing is in your workspace's routing pool - nothing else to configure
curl -s -X POST https://api.opensmartroute.ai/api/v1/route -H 'Authorization: Bearer $OSR_API_KEY' -H 'Content-Type: application/json' -d '{"text": "...", "plan": true}'
# or pin it on the OpenAI-compatible endpoint: {"model": "model-nvidia-nemotron-3-ultra-550b-a55b-free", ...}
Manifest
An Open Capability Manifest: the router reads it to know what this does, what it costs and when to pick it.