Compare models
Put models side by side on price, context, capabilities, and benchmarks.
DeepSeek V3.1 Terminus is a refined checkpoint of DeepSeek V3.1, tuned for improved stability, agentic tool use, and coding performance. It is positioned as a polished general-purpose and agentic workhorse ahead of the V3.2/V4 generations.
Legal & Government#62Medicine & Healthcare#75Life, Physical, & Social Science#90Writing, Literature, & Language#91
Cost rate
0.2x
Context
164K
Input
Output
Reasoning
Tool calling
Structured outputs
Pricing/M tokens
Input$0.27
Cached input$0.14
Output$1.00
Context
Context length163,840
Max output tokens32,768
Knowledge cutoff2025-03-31
Performance
Throughput-
Latency-
Claude Haiku 4.5 is Anthropic's fast, affordable small model in the Claude 4 generation, balancing strong instruction-following and coding ability with low latency. It is designed for high-throughput agentic and chat applications where speed and cost efficiency are priorities.
Document analysis#31Software & IT Services#100Mathematical#101Writing, Literature, & Language#109
Cost rate
1x
Context
200K
Input
Output
Reasoning
Tool calling
Structured outputs
Pricing/M tokens
Input$1.00
Cached input$0.10
Output$5.00
Context
Context length200,000
Max output tokens64,000
Knowledge cutoff-
Performance
Throughput90 tok/s
Latency1.04s