Compare models
Put models side by side on price, context, capabilities, and benchmarks.
Batch-processing variant of Claude Haiku 4.5, Anthropic's fast and affordable Claude 4 small model. Same capabilities as the standard model, offered at lower cost via asynchronous batch processing.
Document analysis#31Mathematical#100Software & IT Services#101Writing, Literature, & Language#110
AuthorAnthropic
Country🇺🇸 United States
ReleasedOct 15, 2025
Overview
Cost rate
0.5x
Context
200K
Input
Output
Reasoning
Tool calling
Structured outputs
Pricing/M tokens
Input$0.50
Cached input$0.05
Output$2.50
Context
Context length200,000
Max output tokens64,000
Knowledge cutoff-
Performance
Throughput-
Latency-
GPT-5.6 Luna is the fast, low-cost tier of OpenAI's GPT-5.6 family, tuned for high-volume, latency-sensitive work like chat, classification, and lightweight agents while retaining genuine GPT-5.6 reasoning.
Document analysis#19Business, Management, & Finance#35Mathematical#37Image understanding#38
AuthorOpenAI
Country🇺🇸 United States
ReleasedJul 9, 2026
Overview
Cost rate
0.2x
Context
1.1M
Input
Output
Reasoning
Tool calling
Structured outputs
Pricing/M tokens
Input$0.20
Cached input$0.02
Output$1.20
Context
Context length1,050,000
Max output tokens128,000
Knowledge cutoff2026-02-16
Performance
Throughput134 tok/s
Latency2.22s
Benchmarks
Head-to-head scores across reasoning, coding, and agentic indices, plus per-domain rankings.
Intelligence
62.5
Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
62.1
Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback)
59.5