Qwen3 VL 8B Instruct
Qwen · Open weights · Released Oct 14, 2025 · qwen/qwen3-vl-8b-instruct
Best $0.117 in / $0.455 out via Alibaba · 131K ctx at this price (headline 262K)
262K contextVisionToolsStructured outputPrompt cachingOpen weights
Available from 2 providersUSD / MTOK · PRICES 23 MINUTES AGO · SPEED 17 DAYS AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $0.117 | $0.455 | — | 131K | — | 1.4s | 121 tok/s | 100.0% | |
| $0.25 | $0.75 | $0.12 | 262K | bf16 | 1.2s | 70 tok/s | 96.9% |
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · methodology v1.1 · measured 17 days ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
Recent changesRSS ↗
Oct 5, 2026endpoint removed · primeintellect3 → 2
Oct 4, 2026endpoint added · primeintellect2 → 3
About
Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...
Specifications
Context window262,144
Max output32,768
Modalities inimage, text
Modalities outtext
LicenseOpen weights
ReleasedOct 14, 2025
TokenizerQwen3
WeightsHugging Face ↗
Open weights
safetensors
Related
Embed badge
[](https://modelindex.ai/models/qwen/qwen3-vl-8b-instruct)Sources
OpenRouter API23 minutes ago
AGGREGATORHugging Face23 minutes ago
COMMUNITYModelIndex benchmarks19 days ago
MEASUREDVerified 23 minutes ago · list prices, not negotiated rates.