Models/Qwen/Qwen3 VL 32B Instruct

Qwen3 VL 32B Instruct

Qwen · Open weights · Released Oct 23, 2025 · qwen/qwen3-vl-32b-instruct
CompareEstimate cost
131K contextVisionToolsStructured outputOpen weights
Available from 1 providerUSD / MTOK · VERIFIED 4 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
PROVIDERINPUTOUTPUTCACHECONTEXTQUANTTTFTTPSUPTIME
TTFT/TPS: median (P50) over trailing 72h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · methodology v1.1 · measured 43 hours ago
About

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

Specifications
Context window131,072
Max output32,768
Modalities intext, image
Modalities outtext
LicenseOpen weights
ReleasedOct 23, 2025
TokenizerQwen
Open weights
Downloads879.8K
Likes234
Licenseapache-2.0
safetensors
Embed badge
Qwen3 VL 32B Instruct price and speed badge[![Qwen3 VL 32B Instruct](https://modelindex.ai/badge/qwen/qwen3-vl-32b-instruct)](https://modelindex.ai/models/qwen/qwen3-vl-32b-instruct)
Sources
OpenRouter API4 minutes ago
AGGREGATOR
Hugging Face4 minutes ago
COMMUNITY
ModelIndex benchmarks43 hours ago
MEASURED
Verified 4 minutes ago · list prices, not negotiated rates.