Qwen3 VL 32B Instruct
Qwen · Open weights · Released Oct 23, 2025 · qwen/qwen3-vl-32b-instruct
131K contextVisionToolsStructured outputOpen weights
Available from 1 providerUSD / MTOK · VERIFIED 4 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $0.104 | $0.416 | — | 131K | — | 2.2s | 77 tok/s | 100.0% |
TTFT/TPS: median (P50) over trailing 72h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · methodology v1.1 · measured 43 hours ago
About
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...
Specifications
Context window131,072
Max output32,768
Modalities intext, image
Modalities outtext
LicenseOpen weights
ReleasedOct 23, 2025
TokenizerQwen
WeightsHugging Face ↗
Open weights
safetensors
Related
Embed badge
[](https://modelindex.ai/models/qwen/qwen3-vl-32b-instruct)Sources
OpenRouter API4 minutes ago
AGGREGATORHugging Face4 minutes ago
COMMUNITYModelIndex benchmarks43 hours ago
MEASUREDVerified 4 minutes ago · list prices, not negotiated rates.