Models/InclusionAI/Ling-3.0-flash

Ling-3.0-flash

InclusionAI · Open weights · Released Jul 23, 2026 · inclusionai/ling-3.0-flash
CompareEstimate cost
262K contextToolsReasoningStructured outputPrompt cachingOpen weights
Available from 3 providersUSD / MTOK · VERIFIED 4 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
PROVIDERINPUTOUTPUTCACHECONTEXTQUANTTTFTTPSUPTIME
TTFT/TPS: median (P50) over trailing 72h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · methodology v1.1 · measured 44 hours ago
Pricing variants · USD / MTok in / out
STANDARD
$0.021 / $0.063
per MTok in / out
CACHE READ $0.004
About

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

Specifications
Context window262,144
Max output32,768
Modalities intext
Modalities outtext
LicenseOpen weights
ReleasedJul 23, 2026
TokenizerOther
Open weights
Downloads17.4K
Likes365
Licensemit
safetensors
Embed badge
Ling-3.0-flash price and speed badge[![Ling-3.0-flash](https://modelindex.ai/badge/inclusionai/ling-3.0-flash)](https://modelindex.ai/models/inclusionai/ling-3.0-flash)
Sources
OpenRouter API4 minutes ago
AGGREGATOR
Hugging Face4 minutes ago
COMMUNITY
ModelIndex benchmarks44 hours ago
MEASURED
Verified 4 minutes ago · list prices, not negotiated rates.