Compare models
Put models side by side on price, context, capabilities, and benchmarks.
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Entertainment, Sports, & Media#225Legal & Government#229Writing, Literature, & Language#239Mathematical#244
AuthorMistral
Country🇫🇷 France
ReleasedDec 1, 2025
Overview
Cost rate
0.3x
Context
262K
Input
Output
Reasoning
Tool calling
Structured outputs
Pricing/M tokens
Input$0.50
Cached input$0.05
Output$1.50
Context
Context length262,144
Max output tokens209,715
Knowledge cutoff-
Performance
Throughput-
Latency-
Claude Haiku 4.5 is Anthropic's fast, affordable small model in the Claude 4 generation, balancing strong instruction-following and coding ability with low latency. It is designed for high-throughput agentic and chat applications where speed and cost efficiency are priorities.
Document analysis#31Mathematical#102Software & IT Services#103Writing, Literature, & Language#110
AuthorAnthropic
Country🇺🇸 United States
ReleasedOct 15, 2025
Overview
Cost rate
1x
Context
200K
Input
Output
Reasoning
Tool calling
Structured outputs
Pricing/M tokens
Input$1.00
Cached input$0.10
Output$5.00
Context
Context length200,000
Max output tokens64,000
Knowledge cutoff-
Performance
Throughput100 tok/s
Latency0.82s
Benchmarks
Head-to-head scores across reasoning, coding, and agentic indices, plus per-domain rankings.
Intelligence
62.5
Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
62.1
Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback)
59.5