Models/Google/Gemini 3.5 Flash

Gemini 3.5 Flash

Google · Proprietary · Released May 19, 2026 · google/gemini-3.5-flash
CompareEstimate cost
1.05M contextVisionToolsReasoningStructured outputPrompt cachingAudio input
Available from 7 providersUSD / MTOK · VERIFIED 12 HOURS AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
PROVIDERINPUTOUTPUTCACHECONTEXTQUANTTTFTTPSUPTIME
TTFT/TPS: median (P50) over trailing 72h · chat10k profile (~10K in / 200 out) · measured against the provider API directly; ° = routed via OpenRouter (adds a hop) · methodology v1.1 · measured 42 hours ago
Pricing variants · USD / MTok in / out
STANDARD
$0.75 / $4.50
per MTok in / out
BATCH
$0.75 / $4.50
async batch
CACHE READ $0.075CACHE WRITE $0.042WEB SEARCH $0.014 / search
About

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

Specifications
Context window1,048,576
Max output65,536
Modalities intext, image, video, file, audio
Modalities outtext
LicenseProprietary
ReleasedMay 19, 2026
Knowledge cutoff2025-01-01
TokenizerGemini
Deprecation2027-05-19
Embed badge
Gemini 3.5 Flash price and speed badge[![Gemini 3.5 Flash](https://modelindex.ai/badge/google/gemini-3.5-flash)](https://modelindex.ai/models/google/gemini-3.5-flash)
Sources
OpenRouter API12 hours ago
AGGREGATOR
LiteLLM dataset12 hours ago
AGGREGATOR
ModelIndex benchmarks42 hours ago
MEASURED
Verified 12 hours ago · list prices, not negotiated rates.