Qwen3 14B
Qwen · Open weights · Released Apr 28, 2025 · qwen/qwen3-14b
Best $0.10 in / $0.22 out via NextBit · int4 · 41K ctx at this price (headline 131K)
131K contextToolsReasoningStructured outputOpen weights
Available from 3 providersUSD / MTOK · PRICES 23 MINUTES AGO · SPEED 6 DAYS AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $0.10 | $0.22 | — | 41K | int4 | 4.4s | 66 tok/s | 99.9% | |
| $0.12 | $0.24 | — | 41K | fp8 | 0.86s | 57 tok/s | 99.4% | |
| $0.228 | $0.91 | — | 131K | — | ✕ | ✕ | 100.0% |
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · ✕ = probed, endpoint returned no measurable stream · methodology v1.1 · measured 6 days ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
About
Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
Specifications
Context window131,072
Max output16,384
Modalities intext
Modalities outtext
LicenseOpen weights
ReleasedApr 28, 2025
Knowledge cutoff2025-03-31
TokenizerQwen3
WeightsHugging Face ↗
Related
Embed badge
[](https://modelindex.ai/models/qwen/qwen3-14b)Sources
OpenRouter API23 minutes ago
AGGREGATORHugging Face23 minutes ago
COMMUNITYModelIndex benchmarks20 days ago
MEASUREDVerified 23 minutes ago · list prices, not negotiated rates.