DeepSeek V4.1 Flash vs gpt-oss-120b
API pricing, context, and independently measured performance, side by side. gpt-oss-120b starts 77% cheaper on input than DeepSeek V4.1 Flash at list prices.
DeepSeek V4.1 Flash
DeepSeek · released Sep 10, 2026
Best input$0.13 / MTok
Best output$0.42 / MTok
Context1.05M
Providers22
Fastest measured0.64s° TTFT · 281 tok/s
Open weightsVisionToolsReasoningStructured outputPrompt caching
gpt-oss-120b
OpenAI · released Aug 5, 2025
Best input$0.03 / MTok
Best output$0.17 / MTok
Context131K
Providers24
Fastest measured0.28s° TTFT · 226 tok/s
Open weightsToolsReasoningStructured outputPrompt caching
Cheapest providers, measuredUSD / MTOK
| MODEL | PROVIDER | INPUT | OUTPUT | TTFT | TPS | SMOKE |
|---|---|---|---|---|---|---|
| DeepSeek V4.1 Flash | RERelace | $0.13 | $0.52 | — | — | — |
| DeepSeek V4.1 Flash | MOMorph | $0.135 | $0.54 | — | — | — |
| DeepSeek V4.1 Flash | $0.14 | $0.42 | — | — | — | |
| gpt-oss-120b | AKAkashML | $0.03 | $0.17 | 1.3s° | 71° tok/s | 4/5 |
| gpt-oss-120b | CWCoreWeave | $0.03 | $0.17 | 0.69s° | 96° tok/s | 3/5 |
| gpt-oss-120b | DEDekaLLM | $0.03 | $0.18 | — | — | — |
TTFT/TPS: ModelIndex-measured medians (P50, trailing 21d) · chat10k profile (~10K in / 200 out) · ° = measured via OpenRouter routing (adds a hop) · SMOKE: of 5 programmatic capability probes passed · methodology v1.1 — full details on each model page