Compare models
Put models side by side on price, context, capabilities, and benchmarks.
DeepSeek R1 Distill Llama 70B is a Llama 3.3 70B model distilled from DeepSeek R1's reasoning traces, transferring much of R1's chain-of-thought reasoning ability into a smaller, faster base architecture. It suits math, coding, and logic tasks where full R1 would be too costly or slow.
Claude Haiku 4.5 is Anthropic's fast, affordable small model in the Claude 4 generation, balancing strong instruction-following and coding ability with low latency. It is designed for high-throughput agentic and chat applications where speed and cost efficiency are priorities.
Benchmarks
Head-to-head scores across reasoning, coding, and agentic indices, plus per-domain rankings.