Models/Google/Gemini 2.5 Flash

Gemini 2.5 Flash

Google · Proprietary · Released Jun 17, 2025 · google/gemini-2.5-flash
Best $0.30 in / $2.50 out via Google
CompareEstimate cost
1.05M contextVisionToolsReasoningStructured outputPrompt cachingAudio input
Available from 9 providersUSD / MTOK · PRICES 23 MINUTES AGO · SPEED 8 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
PROVIDERINPUTOUTPUTCACHECONTEXTQUANTTTFTTPSUPTIME
TTFT/TPS: median (P50) over trailing 504h · chat10k profile (~10K in / 200 out) · measured against the provider API directly; ° = routed via OpenRouter (adds a hop) · ✕ = probed, endpoint returned no measurable stream · methodology v1.1 · measured 8 minutes ago
SERVE THIS MODEL AND NOT LISTED? LIST YOUR INFERENCE →
Pricing variants · USD / MTok in / out
STANDARD
$0.30 / $2.50
per MTok in / out
BATCH
$0.15 / $1.25
async batch
CACHE READ $0.03CACHE WRITE $0.083WEB SEARCH $0.014 / search
Recent changesRSS ↗
Sep 23, 2026endpoint added · cheaperinference7 → 9
About

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

Specifications
Context window1,048,576
Max output65,535
Modalities infile, image, text, audio, video
Modalities outtext
LicenseProprietary
ReleasedJun 17, 2025
Knowledge cutoff2025-01-31
TokenizerGemini
Deprecation2026-10-20
Embed badge
Gemini 2.5 Flash price and speed badge
[![Gemini 2.5 Flash](https://modelindex.ai/badge/google/gemini-2.5-flash)](https://modelindex.ai/models/google/gemini-2.5-flash)
Sources
OpenRouter API23 minutes ago
AGGREGATOR
LiteLLM dataset23 minutes ago
AGGREGATOR
ModelIndex benchmarks18 days ago
MEASURED
Verified 23 minutes ago · list prices, not negotiated rates.