Quality
How Meta: Llama 3.1 8B Instruct compares against every tracked model
Composite of GPQA Diamond, Mock AIME 2024–25, MATH Level 5, SWE-bench Verified, LiveBench, Humanity's Last Exam and Terminal-Bench, normalised 0–100. Scores are ingested from Epoch AI; we do not run these evaluations.
Quality Breakdown
Meta: Llama 3.1 8B Instruct on each evaluation feeding the index, against the tracked average
Bar = Meta: Llama 3.1 8B Instruct · grey marker = average across tracked models · scores ingested from Epoch AI
Arena Elo
Human preference rating from blind pairwise votes
Elo from blind pairwise votes, ingested from Arena
Ingested 29 Jul 2026.
Price & Cost
Published API pricing by token type
USD per 1M tokens at a 3:1 input:output blend. Live from OpenRouter, cross-checked against models.dev. Lower is better.
Context Window
Maximum input tokens accepted
Maximum input context from provider metadata via models.dev and OpenRouter. Higher is better.
Instructthis model #293/375
Speed
Median output throughput across hosting providers
Median output tokens per second across all tracked providers. Derived from OpenRouter endpoint telemetry. Higher is better.
Instructthis model #117/317
Latency
Time to first token, median across providers
Seconds to first streamed chunk, median across tracked providers. Derived from OpenRouter endpoint telemetry. Lower is better.
Instructthis model #37/317 · off scale
Providers
Hosts serving Meta: Llama 3.1 8B Instruct, with their own price and speed
| Provider | Input $/M | Output $/M | Blended | Speed | TTFT | Context |
|---|---|---|---|---|---|---|
| Groq | $0.05 | $0.08 | $0.06 | 109 | 0.3s | 131k |
| CoreWeave | $0.22 | $0.22 | $0.22 | 100 | 0.2s | 128k |
| Novita | $0.02 | $0.05 | $0.03 | 83 | 0.5s | 16k |
| DeepInfra | $0.02 | $0.04 | $0.03 | 20 | 0.6s | 131k |
| Cloudflare | $0.15 | $0.29 | $0.19 | 18 | 0.6s | 32k |
Per-host pricing and telemetry from OpenRouter. Speed is output tokens per second; TTFT is time to first token.
History
How price and quality have moved since we began tracking this model
Since 25 Jul 2026
| Known Good Index | — |
| Blended price | $0.06 per 1M · observed 29 Jul 2026 |
| Output speed | 59 tok/s · observed 29 Jul 2026 |
| Time to first token | 0.4s |
No recorded changes for this model since tracking began.
Comparisons
Head-to-head pages pairing Meta: Llama 3.1 8B Instruct with other tracked models
Head-to-head pages are generated for the top 50 models by Known Good Index, which Meta: Llama 3.1 8B Instruct is not currently in. Its nearest tracked models by index are below.