Compare models
Put models side by side on price, context, capabilities, and benchmarks.
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks.
Mathematical#4Medicine & Healthcare#8Life, Physical, & Social Science#12Software & IT Services#13
AuthorZ.AI
Countryπ¨π³ China
ReleasedAug 18, 2026
Overview
Cost rate
0.5x
Context
1M
Input
Output
Reasoning
Tool calling
Structured outputs
Pricing/M tokens
Input$0.70
Cached input$0.13
Output$2.20
Context
Context length1,048,576
Max output tokens943,718
Knowledge cutoff-
Performance
Throughput-
Latency-
Claude Haiku 4.5 is Anthropic's fast, affordable small model in the Claude 4 generation, balancing strong instruction-following and coding ability with low latency. It is designed for high-throughput agentic and chat applications where speed and cost efficiency are priorities.
Document analysis#31Mathematical#103Software & IT Services#104Writing, Literature, & Language#112
AuthorAnthropic
CountryπΊπΈ United States
ReleasedOct 15, 2025
Overview
Cost rate
1x
Context
200K
Input
Output
Reasoning
Tool calling
Structured outputs
Pricing/M tokens
Input$1.00
Cached input$0.10
Output$5.00
Context
Context length200,000
Max output tokens64,000
Knowledge cutoff-
Performance
Throughput81 tok/s
Latency0.78s
Benchmarks
Head-to-head scores across reasoning, coding, and agentic indices, plus per-domain rankings.
Intelligence
53.4
Claude Fable 5.1 (Adaptive Reasoning, Max Effort, Default Fallback)
52.8
GPT-6 Astra (max)
50.7