Compare models
Put models side by side on price, context, capabilities, and benchmarks.
The Auto Router automatically selects the best model for your prompt, powered by the wisdom of the market.
AuthorOpenRouter
CountryπΊπΈ United States
ReleasedNov 8, 2023
Overview
Cost rate
-
Context
2M
Input
Output
Reasoning
Tool calling
Structured outputs
Pricing/M tokens
Input-
Cached input-
Output-
Context
Context length2,000,000
Max output tokens-
Knowledge cutoff-
Performance
Throughput-
Latency-
Claude Haiku 4.5 is Anthropic's fast, affordable small model in the Claude 4 generation, balancing strong instruction-following and coding ability with low latency. It is designed for high-throughput agentic and chat applications where speed and cost efficiency are priorities.
Document analysis#31Mathematical#103Software & IT Services#104Writing, Literature, & Language#112
AuthorAnthropic
CountryπΊπΈ United States
ReleasedOct 15, 2025
Overview
Cost rate
1x
Context
200K
Input
Output
Reasoning
Tool calling
Structured outputs
Pricing/M tokens
Input$1.00
Cached input$0.10
Output$5.00
Context
Context length200,000
Max output tokens64,000
Knowledge cutoff-
Performance
Throughput86 tok/s
Latency0.75s
Benchmarks
Head-to-head scores across reasoning, coding, and agentic indices, plus per-domain rankings.
Intelligence
62.5
Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
62.1
Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback)
59.5