Quality
How DeepSeek: R1 Distill Llama 70B compares against every tracked model
Composite of GPQA Diamond, Mock AIME 2024–25, MATH Level 5, SWE-bench Verified, LiveBench, Humanity's Last Exam and Terminal-Bench, normalised 0–100. Scores are ingested from Epoch AI, LiveBench; we do not run these evaluations.
Quality Breakdown
DeepSeek: R1 Distill Llama 70B on each evaluation feeding the index, against the tracked average
Bar = DeepSeek: R1 Distill Llama 70B · grey marker = average across tracked models · scores ingested from Epoch AI, LiveBench
Arena Elo
Human preference rating from blind pairwise votes
Elo from blind pairwise votes, ingested from Arena
DeepSeek: R1 Distill Llama 70B does not currently appear in any arena we ingest.
Price & Cost
Published API pricing by token type
USD per 1M tokens at a 3:1 input:output blend. Live from OpenRouter, cross-checked against models.dev. Lower is better.
70Bthis model #187/370 · off scale
Context Window
Maximum input tokens accepted
Maximum input context from provider metadata via models.dev and OpenRouter. Higher is better.
70Bthis model #368/375
Speed
Median output throughput across hosting providers
Median output tokens per second across all tracked providers. Derived from OpenRouter endpoint telemetry. Higher is better.
70Bthis model #268/317
Latency
Time to first token, median across providers
Seconds to first streamed chunk, median across tracked providers. Derived from OpenRouter endpoint telemetry. Lower is better.
70Bthis model #97/317 · off scale
Providers
Hosts serving DeepSeek: R1 Distill Llama 70B, with their own price and speed
| Provider | Input $/M | Output $/M | Blended | Speed | TTFT | Context |
|---|---|---|---|---|---|---|
| Novita | $0.80 | $0.80 | $0.80 | 32 | 0.8s | 8k |
Per-host pricing and telemetry from OpenRouter. Speed is output tokens per second; TTFT is time to first token.
History
How price and quality have moved since we began tracking this model
Since 25 Jul 2026
| Known Good Index | 62 · based on 4 of 7 evaluations |
| Blended price | $0.80 per 1M · observed 29 Jul 2026 |
| Output speed | 25 tok/s · observed 29 Jul 2026 |
| Time to first token | 0.8s |
No recorded changes for this model since tracking began.
Comparisons
Head-to-head pages pairing DeepSeek: R1 Distill Llama 70B with other tracked models