Hi - I answer from the OpenSmartRoute documentation: routing, the API, plans and quotas, self-hosting. Ask away, or open a support ticket if you need a person.
Grounded in the docs - follow a source before acting on it.
Qwen3.8 2.4T A95B (batch) - Model - OpenSmartRoute
Modelv1.0.0
Qwen3.8 2.4T A95B (batch)
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion tot
Reference model card from the OpenRouter listing - model qwen/qwen3.8-2.4t-a95b:batch. Prices are per million tokens as listed; ratings and installs below are from this marketplace.
Qwen3.8 2.4T A95B (batch)
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of Qwen3.8 Max, with 95 billion active parameters out of 2.4 trillion total. It is...
Facts
Vendor
Qwen
Context window
1,010,000 tokens
Max output
909,000 tokens
Input price
$2.00 / 1M tokens
Output price
$6.00 / 1M tokens
Input
text
Output
text
Tool calling
yes
Structured outputs
yes
Reasoning
yes
Open weights
Qwen/Qwen3.8-2.4T-A95B
Endpoint: https://openrouter.ai/api/v1 (OpenAI-compatible), model qwen/qwen3.8-2.4t-a95b:batch
Use it
Copy one of these into your project. Installing also returns the manifest and these snippets.
# after Install: the listing is in your workspace's routing pool - nothing else to configure
curl -s -X POST https://api.opensmartroute.ai/api/v1/route -H 'Authorization: Bearer $OSR_API_KEY' -H 'Content-Type: application/json' -d '{"text": "...", "plan": true}'
# or pin it on the OpenAI-compatible endpoint: {"model": "model-qwen-qwen3-8-2-4t-a95b-batch", ...}
Manifest
An Open Capability Manifest: the router reads it to know what this does, what it costs and when to pick it.
model-qwen-qwen3-8-2-4t-a95b-batch.ocm.jsonjson
{
"ocm": "1",
"id": "model-qwen-qwen3-8-2-4t-a95b-batch",
"kind": "llm",
"name": "Qwen3.8 2.4T A95B (batch)",
"description": "Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...",
"publisher": "Qwen",
"version": "1.0.0",
"capabilities": {
"domains": [
"general"
],
"tags": [
"llm",
"qwen",
"openrouter",
"open-weights",
"reasoning",
"tool-calling",
"models"
],
"languages": [
"en"
],
"modalities": [
"text"
],
"supports_tools": true,
"supports_streaming": true,
"context_window": 1010000
},
"quality_prior": 0.4,
"examples": [
"Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is..."
],
"primary": false,
"metadata": {
"source": {
"provider": "models",
"ref": "1786551702",
"url": "https://openrouter.ai/qwen/qwen3.8-2.4t-a95b:batch",
"key": "qwen/qwen3.8-2.4t-a95b:batch",
"catalogue": "https://openrouter.ai/api/v1"
},
"model": "qwen/qwen3.8-2.4t-a95b:batch",
"pricing": {
"input_usd_per_1m": 2,
"output_usd_per_1m": 6,
"cache_read_usd_per_1m": 0.25
},
"max_output_tokens": 909000,
"output_modalities": [
"text"
],
"hugging_face_id": "Qwen/Qwen3.8-2.4T-A95B",
"reasoning": true,
"created": 1786551702,
"benchmarks": {
"intelligence_index": 40,
"coding_index": 71.9,
"agentic_index": 50.4,
"osr_index": 43.1,
"osr_coverage": 0.4
}
},
"cost": {
"usd_per_1k_tokens": 0.004
},
"endpoints": [
{
"protocol": "openai",
"url": "https://openrouter.ai/api/v1"
}
]
}
Fetch it by URL: GET /api/v1/registry/model-qwen-qwen3-8-2-4t-a95b-batch/manifest?version=1.0.0
Reviews
Star ratings from people who tried it. One review per account; edit yours any time.
No reviews yet. Install it, try it, and be the first to rate it.