Models/Qwen/Qwen3 VL 32B Instruct

Qwen3 VL 32B Instruct

Qwen · Open weights · Released Oct 23, 2025 · qwen/qwen3-vl-32b-instruct
Best $0.104 in / $0.416 out via Alibaba
CompareEstimate cost
131K contextVisionToolsStructured outputOpen weights
Available from 1 providerUSD / MTOK · PRICES 23 MINUTES AGO · SPEED 8 DAYS AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
PROVIDERINPUTOUTPUTCACHECONTEXTQUANTTTFTTPSUPTIME
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · methodology v1.1 · measured 8 days ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
About

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

Specifications
Context window131,072
Max output32,768
Modalities intext, image
Modalities outtext
LicenseOpen weights
ReleasedOct 23, 2025
TokenizerQwen
Open weights
Downloads322.3K
Likes244
Licenseapache-2.0
safetensors
Embed badge
Qwen3 VL 32B Instruct price and speed badge
[![Qwen3 VL 32B Instruct](https://modelindex.ai/badge/qwen/qwen3-vl-32b-instruct)](https://modelindex.ai/models/qwen/qwen3-vl-32b-instruct)
Sources
OpenRouter API23 minutes ago
AGGREGATOR
Hugging Face23 minutes ago
COMMUNITY
ModelIndex benchmarks8 days ago
MEASURED
Verified 23 minutes ago · list prices, not negotiated rates.