Models/Llama 3.3 70B Instruct vs MiniMax M3

Llama 3.3 70B Instruct vs MiniMax M3

API pricing, context, and independently measured performance, side by side. Llama 3.3 70B Instruct starts 57% cheaper on input than MiniMax M3 at list prices.

Llama 3.3 70B Instruct
Meta · released Dec 6, 2024
Best input$0.10 / MTok
Best output$0.32 / MTok
Context131K
Providers13
Fastest measured0.59s° TTFT · 86 tok/s
Open weightsToolsStructured outputPrompt caching
MiniMax M3
MiniMax · released May 31, 2026
Best input$0.23 / MTok
Best output$0.96 / MTok
Context1.05M
Providers13
Fastest measured0.37s° TTFT · 188 tok/s
Open weightsVisionToolsReasoningStructured outputPrompt caching
Cheapest providers, measuredUSD / MTOK
MODELPROVIDERINPUTOUTPUTTTFTTPSSMOKE
Llama 3.3 70B InstructDeepInfra logoDeepInfra$0.10$0.321.9s°12° tok/s5/5
Llama 3.3 70B InstructNebius logoNebius$0.13$0.4015s°3° tok/s5/5
Llama 3.3 70B InstructNovita logoNovita$0.135$0.401.8s°25° tok/s5/5
MiniMax M3CWCoreWeave$0.23$0.960.93s°98° tok/s5/5
MiniMax M3GMIGMICloud$0.24$0.962.5s°136° tok/s5/5
MiniMax M3MOMorph$0.255$1.022.2s°50° tok/s4/5
TTFT/TPS: ModelIndex-measured medians (P50, trailing 72h) · chat10k profile (~10K in / 200 out) · ° = measured via OpenRouter routing (adds a hop) · SMOKE: of 5 programmatic capability probes passed · methodology v1.1 — full details on each model page
All 13 providers for Llama 3.3 70B InstructAll 13 providers for MiniMax M3Open in comparator