R1 Distill Llama 70B
DeepSeek
deepseek/deepseek-r1-distill-llama-70bDeepSeek R1 Distill Llama 70B is a Llama 3.3 70B model distilled from DeepSeek R1's reasoning traces, transferring much of R1's chain-of-thought reasoning ability into a smaller, faster base architecture. It suits math, coding, and logic tasks where full R1 would be too costly or slow.
Cost rate
0.3x
Context
8K
Released
Jan 23, 2025
Input
Text
Output
Text
Support
ReasoningTool callingStructured outputs
Best at
The categories where this model ranks highest.
No top-ranked categories found.
Performance
Median latency and throughput measured across recent requests.
Throughput25 tok/s
Latency