Llama 3.1 70B Instruct
Meta · Open weights · Released Jul 23, 2024 · meta-llama/llama-3.1-70b-instruct
131K contextToolsStructured outputPrompt cachingOpen weights
Available from 3 providersUSD / MTOK · VERIFIED 4 MINUTES AGO
EST. MONTHLY COSTclick a row for detail · click headers to sort
| PROVIDER | INPUT | OUTPUT | CACHE | CONTEXT | QUANT | TTFT | TPS | UPTIME |
|---|---|---|---|---|---|---|---|---|
| $0.40 | $0.40 | — | 131K | fp8 | 0.95s | 52 tok/s | 99.6% | |
| $0.72 | $0.72 | — | 131K | — | 1.1s | 78 tok/s | 100.0% | |
| CWCoreWeave | $0.80 | $0.80 | $0.80 | 128K | bf16 | 0.61s | 76 tok/s | 100.0% |
TTFT/TPS: median (P50) over trailing 72h · chat10k profile (~10K in / 200 out) · measured via OpenRouter routing · methodology v1.1 · measured 44 hours ago
About
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong...
Specifications
Context window131,072
Max output16,384
Modalities intext
Modalities outtext
LicenseOpen weights
ReleasedJul 23, 2024
Knowledge cutoff2023-12-31
TokenizerLlama3
WeightsHugging Face ↗
Open weightsGATED
safetensorspytorch
Related
Embed badge
[](https://modelindex.ai/models/meta-llama/llama-3.1-70b-instruct)Sources
OpenRouter API4 minutes ago
AGGREGATORHugging Face4 minutes ago
COMMUNITYModelIndex benchmarks44 hours ago
MEASUREDVerified 4 minutes ago · list prices, not negotiated rates.