Qwen2.5 VL 72B Instruct
Qwen · Open weights · Released Feb 1, 2025 · qwen/qwen2.5-vl-72b-instruct
Best $0.80 in / $1.00 out via Parasail · fp8
128K contextVisionStructured outputPrompt cachingOpen weights
Available from 1 providerUSD / MTOK · PRICES 23 MINUTES AGO · SPEED 2 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $0.80 | $1.00 | $0.40 | 128K | fp8 | 1.4s | 46 tok/s | 100.0% |
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · methodology v1.1 · measured 2 minutes ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
Pricing variants · USD / MTok in / out
STANDARD
$0.80 / $1.00
per MTok in / out
CACHE READ $0.40
Recent changesRSS ↗
Sep 4, 2026endpoint removed · nebius2 → 1
About
Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.
Specifications
Context window128,000
Max output115,200
Modalities intext, image
Modalities outtext
LicenseOpen weights
ReleasedFeb 1, 2025
Knowledge cutoff2024-06-30
TokenizerQwen
WeightsHugging Face ↗
Related
Embed badge
[](https://modelindex.ai/models/qwen/qwen2.5-vl-72b-instruct)Sources
OpenRouter API23 minutes ago
AGGREGATORHugging Face23 minutes ago
COMMUNITYModelIndex benchmarks2 minutes ago
MEASUREDVerified 23 minutes ago · list prices, not negotiated rates.