modelgrep
deepseek logo

DeepSeek: R1 Distill Llama 70B

deepseek/deepseek-r1-distill-llama-70b

154th smartest of 182Reasoning
Use via OpenRouter ↗
Intelligence
9.9
154th of 182
Design Elo
Speed
tokens/sec
Latency
first token
Input price
$0.800
211th cheapest
Context
8K
8K max out

How it compares

Smarter than15%
of all ranked models
Cheaper than36%
of all ranked models

Overview

DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...

Benchmarks

independent · Artificial Analysis & Design Arena
Artificial Analysis36th percentile
Intelligence Index
9.9
GPQA Diamond
40%
Humanity's Last Exam
6%
SciCode
31%
Tau²-Bench (agentic)
22%

Providers & pricing (1)

ProviderIn $/MOut $/MUptime
Novitabf16$0.800$0.800100%

Specifications

Context window8K
Max output8K
Knowledge cutoffJul 2024
Input modalitiestext
Output modalitiestext
Prompt caching
Cache read price
ModeratedNo

R1 Distill Llama 70B FAQ

How much does R1 Distill Llama 70B cost?

R1 Distill Llama 70B costs $0.800 per million input tokens and $0.800 per million output tokens via OpenRouter, making it 211th cheapest of 332 paid models.

How smart is R1 Distill Llama 70B?

R1 Distill Llama 70B scores 9.9 on the Artificial Analysis Intelligence Index, ranking 154th of 182 benchmarked models, with a GPQA Diamond score of 40%.

What is R1 Distill Llama 70B's context window?

R1 Distill Llama 70B supports a 8K-token context window and can output up to 8K tokens. It accepts text input.

Compare head-to-head