Quality
How Meta: Llama 3.3 70B Instruct compares against every tracked model
Composite of GPQA Diamond, Mock AIME 2024–25, MATH Level 5, SWE-bench Verified, LiveBench, Humanity's Last Exam and Terminal-Bench, normalised 0–100. Scores are ingested from Epoch AI, LiveBench; we do not run these evaluations.
Instructthis model #53/62
Quality Breakdown
Meta: Llama 3.3 70B Instruct on each evaluation feeding the index, against the tracked average
Bar = Meta: Llama 3.3 70B Instruct · grey marker = average across tracked models · scores ingested from Epoch AI, LiveBench
Arena Elo
Human preference rating from blind pairwise votes
Elo from blind pairwise votes, ingested from Arena
Ingested 29 Jul 2026.
Price & Cost
Published API pricing by token type
USD per 1M tokens at a 3:1 input:output blend. Live from OpenRouter, cross-checked against models.dev. Lower is better.
Instructthis model #74/370 · off scale
Context Window
Maximum input tokens accepted
Maximum input context from provider metadata via models.dev and OpenRouter. Higher is better.
Instructthis model #279/375
Speed
Median output throughput across hosting providers
Median output tokens per second across all tracked providers. Derived from OpenRouter endpoint telemetry. Higher is better.
Instructthis model #187/317
Latency
Time to first token, median across providers
Seconds to first streamed chunk, median across tracked providers. Derived from OpenRouter endpoint telemetry. Lower is better.
Instructthis model #57/317 · off scale
Providers
Hosts serving Meta: Llama 3.3 70B Instruct, with their own price and speed
| Provider | Input $/M | Output $/M | Blended | Speed | TTFT | Context |
|---|---|---|---|---|---|---|
| Groq | $0.59 | $0.79 | $0.64 | 174 | 0.3s | 131k |
| Google · first party | $0.72 | $0.72 | $0.72 | 72 | 0.3s | 128k |
| Crusoe | $0.25 | $0.75 | $0.38 | 68 | 0.3s | 131k |
| SambaNova | $0.45 | $0.90 | $0.56 | 65 | 0.7s | 16k |
| CoreWeave | $0.71 | $0.71 | $0.71 | 52 | 0.3s | 128k |
| Parasail | $0.22 | $0.50 | $0.29 | 36 | 0.6s | 131k |
| Novita | $0.14 | $0.40 | $0.20 | 31 | 0.6s | 6k |
| Cloudflare | $0.29 | $2.25 | $0.78 | 24 | 0.7s | 24k |
| Nebius | $0.13 | $0.40 | $0.20 | 24 | 0.6s | 131k |
| Together | $1.04 | $1.04 | $1.04 | 21 | 1.1s | 131k |
| AkashML | $0.13 | $0.40 | $0.20 | 20 | 0.7s | 131k |
| DeepInfra | $0.10 | $0.32 | $0.15 | 13 | 0.7s | 131k |
Per-host pricing and telemetry from OpenRouter. Speed is output tokens per second; TTFT is time to first token.
History
How price and quality have moved since we began tracking this model
Since 25 Jul 2026
| Known Good Index | 34 · based on 4 of 7 evaluations |
| Blended price | $0.20 per 1M · observed 29 Jul 2026 |
| Output speed | 44 tok/s · observed 29 Jul 2026 |
| Time to first token | 0.6s |
Comparisons
Head-to-head pages pairing Meta: Llama 3.3 70B Instruct with other tracked models
Head-to-head pages are generated for the top 50 models by Known Good Index, which Meta: Llama 3.3 70B Instruct is not currently in. Its nearest tracked models by index are below.