Hi - I answer from the OpenSmartRoute documentation: routing, the API, plans and quotas, self-hosting. Ask away, or open a support ticket if you need a person.
Grounded in the docs - follow a source before acting on it.
Qwen3 VL 8B Thinking - Model - OpenSmartRoute
Modelv1.0.0
Qwen3 VL 8B Thinking
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences
Reference model card from the OpenRouter listing - model qwen/qwen3-vl-8b-thinking. Prices are per million tokens as listed; ratings and installs below are from this marketplace.
Qwen3 VL 8B Thinking
Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...
Facts
Vendor
Qwen
Context window
131,072 tokens
Max output
32,768 tokens
Input price
$0.18 / 1M tokens
Output price
$2.10 / 1M tokens
Input
image, text
Output
text
Tool calling
yes
Structured outputs
yes
Reasoning
yes
Open weights
Qwen/Qwen3-VL-8B-Thinking
Endpoint: https://openrouter.ai/api/v1 (OpenAI-compatible), model qwen/qwen3-vl-8b-thinking
Use it
Copy one of these into your project. Installing also returns the manifest and these snippets.
# after Install: the listing is in your workspace's routing pool - nothing else to configure
curl -s -X POST https://api.opensmartroute.ai/api/v1/route -H 'Authorization: Bearer $OSR_API_KEY' -H 'Content-Type: application/json' -d '{"text": "...", "plan": true}'
# or pin it on the OpenAI-compatible endpoint: {"model": "model-qwen-qwen3-vl-8b-thinking", ...}
Manifest
An Open Capability Manifest: the router reads it to know what this does, what it costs and when to pick it.
model-qwen-qwen3-vl-8b-thinking.ocm.jsonjson
{
"ocm": "1",
"id": "model-qwen-qwen3-vl-8b-thinking",
"kind": "llm",
"name": "Qwen3 VL 8B Thinking",
"description": "Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and...",
"publisher": "Qwen",
"version": "1.0.0",
"capabilities": {
"domains": [
"general"
],
"tags": [
"llm",
"qwen",
"openrouter",
"open-weights",
"reasoning",
"tool-calling",
"image",
"models"
],
"languages": [
"en"
],
"modalities": [
"image",
"text"
],
"supports_tools": true,
"supports_streaming": true,
"context_window": 131072
},
"quality_prior": 0.6,
"examples": [
"Qwen3-VL-8B-Thinking is the reasoning-optimized variant of the Qwen3-VL-8B multimodal model, designed for advanced visual and textual reasoning across complex scenes, documents, and temporal sequences. It integrates enhanced multimodal alignment and..."
],
"primary": false,
"metadata": {
"source": {
"provider": "models",
"ref": "1760463746",
"url": "https://openrouter.ai/qwen/qwen3-vl-8b-thinking",
"key": "qwen/qwen3-vl-8b-thinking",
"catalogue": "https://openrouter.ai/api/v1"
},
"model": "qwen/qwen3-vl-8b-thinking",
"pricing": {
"input_usd_per_1m": 0.18,
"output_usd_per_1m": 2.1
},
"max_output_tokens": 32768,
"output_modalities": [
"text"
],
"hugging_face_id": "Qwen/Qwen3-VL-8B-Thinking",
"reasoning": true,
"created": 1760463746,
"benchmarks": {
"hub_downloads": 161436,
"hub_downloads_total": 2554953,
"hub_likes": 224,
"hub_trending": 0
}
},
"cost": {
"usd_per_1k_tokens": 0.00114
},
"endpoints": [
{
"protocol": "openai",
"url": "https://openrouter.ai/api/v1"
}
]
}
Fetch it by URL: GET /api/v1/registry/model-qwen-qwen3-vl-8b-thinking/manifest?version=1.0.0
Reviews
Star ratings from people who tried it. One review per account; edit yours any time.
No reviews yet. Install it, try it, and be the first to rate it.