GPT-4.1 Nano
OpenAI · Proprietary · Released Apr 14, 2025 · openai/gpt-4.1-nano
1.05M contextVisionToolsStructured outputPrompt caching
Available from 3 providersUSD / MTOK · VERIFIED 4 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $0.10 | $0.40 | $0.03 | 1.05M | — | 0.74s | 165 tok/s | 100.0% | |
| $0.10 | $0.40 | $0.025 | 1.05M | — | 0.55s | 108 tok/s | 99.5% | |
| $0.11 | $0.44 | $0.033 | 1.05M | — | — | — | — |
TTFT/TPS: median (P50) over trailing 72h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · methodology v1.1 · measured 44 hours ago
Pricing variants · USD / MTok in / out
STANDARD
$0.10 / $0.40
per MTok in / out
BATCH
$0.05 / $0.20
async batch
CACHE READ $0.03WEB SEARCH $0.01 / search
About
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
Specifications
Context window1,047,576
Max output32,768
Modalities inimage, text, file
Modalities outtext
LicenseProprietary
ReleasedApr 14, 2025
Knowledge cutoff2024-06-30
TokenizerGPT
Deprecation2026-10-23
Related
Embed badge
[](https://modelindex.ai/models/openai/gpt-4.1-nano)Sources
OpenRouter API4 minutes ago
AGGREGATORLiteLLM dataset4 minutes ago
AGGREGATORModelIndex benchmarks44 hours ago
MEASUREDVerified 4 minutes ago · list prices, not negotiated rates.