Skip to content
OpenSmartRoute

Models

Routing targets

17 targets across 7 kinds; 4 answer through the OpenAI-compatible proxy and every one is routable for decisions and plans. Each LLM target shows the model it resolves to. Traffic is measured on this deployment over the last 7 days.

targets
17
executable
4
requests · 7d
18
Vendors behind this deploymentMetaOpenAI

Live quotes of representative prompts against this catalogue, refreshed every minute. Quote your own prompt on the cost estimator.

Everyday assistant
Meta
On-prem private model

Llama 3.3 70B Instruct

Recommended
per request
$0.00013
quality
0.66
latency
1.20 s

no rule matched · domain=0.50 action=0.50 complexity_fit=1.00 prior=0.70 · best_sim=0.06 top3_mean=0.02

Coding and refactoring
OpenAIRecommended
per request
$0.00175
quality
0.73
latency
2.50 s

no rule matched · domain=1.00 action=1.00 complexity_fit=0.97 prior=0.93 · best_sim=0.04 top3_mean=0.03

Analysis and reasoning
OpenAIRecommended
per request
$0.00291
quality
0.78
latency
2.50 s

no rule matched · domain=1.00 action=1.00 complexity_fit=1.00 prior=0.93 · best_sim=0.19 top3_mean=0.10

Summaries and rewriting
Meta
On-prem private model

Llama 3.3 70B Instruct

Recommended
per request
$0.0000553
quality
0.76
latency
1.20 s

no rule matched · domain=1.00 action=1.00 complexity_fit=1.00 prior=0.70 · best_sim=0.16 top3_mean=0.08

Extraction and classification
Meta
On-prem private model

Llama 3.3 70B Instruct

Recommended
per request
$0.0000439
quality
0.86
latency
1.20 s

rule 'pii-stays-onprem' prefers · domain=0.50 action=1.00 complexity_fit=1.00 prior=0.70 · best_sim=0.19 top3_mean=0.09

Sensitive data
Meta
On-prem private model

Llama 3.3 70B Instruct

Recommended
per request
$0.0000678
quality
0.89
latency
1.20 s

rule 'pii-stays-onprem' prefers · domain=1.00 action=1.00 complexity_fit=1.00 prior=0.70 · best_sim=0.12 top3_mean=0.08

17 of 17
TargetKindModelPrice / 1MContextLatencyTokens 7dQuality
MetaOn-prem private modelllm-onpremllmexecLlama 3.3 70B InstructMeta via azure/llm-onprem$1.00-638.4 ms212new0.70
UTranslation skillskill-translateskill-$0.50-20.9 ms12new0.85
UHuman escalation queuehuman-escalationhuman-$500-15.9 ms1new0.90
OpenAIFrontier reasoning modelllm-frontierllmexecGPT-4.1OpenAI via azure/llm-frontier$15200K31.3 ms1new0.93
UCoding agentcoding-agentagent-$8.00-6000 ms0no traffic0.85
OpenAIMid-tier modelllm-midllmexecGPT-4.1 MiniOpenAI via azure/llm-mid$3.00-900 ms0no traffic0.75
OpenAISmall fast modelllm-smallllmexecGPT-4.1 NanoOpenAI via azure/llm-small$0.20-300 ms0no traffic0.55
UFriendly support personapersona-friendly-supportpersona-free-0 ms0no traffic0.75
ULegal counsel personapersona-legal-counselpersona-free-0 ms0no traffic0.80
USenior engineer personapersona-senior-engineerpersona-free-0 ms0no traffic0.85
UResearch agentresearch-agentagent-$6.00-8000 ms0no traffic0.80
USQL skillskill-sqlskill-$1.00-700 ms0no traffic0.80
USummarization skillskill-summarizeskill-$0.80-1200 ms0no traffic0.82
UCustomer support agentsupport-agentagent-$2.00-1500 ms0no traffic0.80
UCalculatortool-calculatortool-free-5 ms0no traffic0.99
UImage generatortool-image-gentool-$40-9000 ms0no traffic0.85
UCustomer notification workflowworkflow-notify-customerworkflow-free-2000 ms0no traffic0.90

Add your own targets by editing targets.yaml; set metadata.model to a catalogue id such as anthropic/claude-sonnet-4 to link a target to its reference model. See the guide or the JSON behind this page.