Quality
How Meta: Llama 3.1 70B Instruct compares against every tracked model
Composite of GPQA Diamond, Mock AIME 2024–25, MATH Level 5, SWE-bench Verified, LiveBench, Humanity's Last Exam and Terminal-Bench, normalised 0–100. Scores are ingested from Epoch AI; we do not run these evaluations.
Quality Breakdown
Meta: Llama 3.1 70B Instruct on each evaluation feeding the index, against the tracked average
Bar = Meta: Llama 3.1 70B Instruct · grey marker = average across tracked models · scores ingested from Epoch AI
Arena Elo
Human preference rating from blind pairwise votes
Elo from blind pairwise votes, ingested from Arena
Ingested 29 Jul 2026.
Price & Cost
Published API pricing by token type
USD per 1M tokens at a 3:1 input:output blend. Live from OpenRouter, cross-checked against models.dev. Lower is better.
Instructthis model #123/370 · off scale
Context Window
Maximum input tokens accepted
Maximum input context from provider metadata via models.dev and OpenRouter. Higher is better.
Instructthis model #292/375
Speed
Median output throughput across hosting providers
Median output tokens per second across all tracked providers. Derived from OpenRouter endpoint telemetry. Higher is better.
Instructthis model #232/317
Latency
Time to first token, median across providers
Seconds to first streamed chunk, median across tracked providers. Derived from OpenRouter endpoint telemetry. Lower is better.
Providers
Hosts serving Meta: Llama 3.1 70B Instruct, with their own price and speed
| Provider | Input $/M | Output $/M | Blended | Speed | TTFT | Context |
|---|---|---|---|---|---|---|
| Amazon Bedrock · first party | $0.72 | $0.72 | $0.72 | 31 | 0.3s | 131k |
| CoreWeave | $0.80 | $0.80 | $0.80 | 30 | 0.2s | 128k |
| DeepInfra | $0.40 | $0.40 | $0.40 | 19 | 0.3s | 131k |
Per-host pricing and telemetry from OpenRouter. Speed is output tokens per second; TTFT is time to first token.
History
How price and quality have moved since we began tracking this model
Since 25 Jul 2026
| Known Good Index | — |
| Blended price | $0.40 per 1M · observed 29 Jul 2026 |
| Output speed | 32 tok/s · observed 29 Jul 2026 |
| Time to first token | 0.3s |
No recorded changes for this model since tracking began.
Comparisons
Head-to-head pages pairing Meta: Llama 3.1 70B Instruct with other tracked models
Head-to-head pages are generated for the top 50 models by Known Good Index, which Meta: Llama 3.1 70B Instruct is not currently in. Its nearest tracked models by index are below.