Compare models
Put models side by side on price, context, capabilities, and benchmarks.
GLM-5.1 is an incremental upgrade to Z.ai's GLM-5 flagship model, released as open source in April 2026 with improved reasoning and agentic performance. It continues GLM-5's focus on complex systems design and large-scale programming tasks.
Writing, Literature, & Language#33Entertainment, Sports, & Media#35Mathematical#35Software & IT Services#36
Cost rate
0.7x
Context
205K
Input
Output
Reasoning
Tool calling
Structured outputs
Pricing/M tokens
Input$0.97
Cached input$0.18
Output$3.04
Context
Context length204,800
Max output tokens128,000
Knowledge cutoff-
Performance
Throughput62 tok/s
Latency1.78s
Claude Haiku 4.5 is Anthropic's fast, affordable small model in the Claude 4 generation, balancing strong instruction-following and coding ability with low latency. It is designed for high-throughput agentic and chat applications where speed and cost efficiency are priorities.
Document analysis#31Software & IT Services#100Mathematical#101Writing, Literature, & Language#109
Cost rate
1x
Context
200K
Input
Output
Reasoning
Tool calling
Structured outputs
Pricing/M tokens
Input$1.00
Cached input$0.10
Output$5.00
Context
Context length200,000
Max output tokens64,000
Knowledge cutoff-
Performance
Throughput90 tok/s
Latency1.04s
Benchmarks
Head-to-head scores across reasoning, coding, and agentic indices, plus per-domain rankings.
Intelligence
62.5
Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
62.5
Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
62.1
Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback)
60.9