Compare models
Put models side by side on price, context, capabilities, and benchmarks.
Kimi K3 is Moonshot AI's July 2026 flagship model, a 2.8-trillion-parameter mixture-of-experts model with native multimodal vision and a 1M-token context window. It is built for long-horizon coding, reasoning, and agent workflows, positioned as an open-weight competitor to top proprietary frontier models.
Legal & Government#2Life, Physical, & Social Science#5Software & IT Services#6Mathematical#6
AuthorMoonshot
Country🇨🇳 China
ReleasedJul 16, 2026
Overview
Cost rate
3x
Context
1M
Input
Output
Reasoning
Tool calling
Structured outputs
Pricing/M tokens
Input$3.00
Cached input$0.30
Output$15.00
Context
Context length1,048,576
Max output tokens1,048,576
Knowledge cutoff-
Performance
Throughput34 tok/s
Latency3.76s
Qwen3.8 Max is Alibaba's proprietary flagship-tier model in the Qwen3.8 generation, intended for top-tier reasoning, coding, and agentic performance alongside the open-weight Qwen3.8 releases.
Image understanding#2Life, Physical, & Social Science#6Medicine & Healthcare#7Entertainment, Sports, & Media#12
AuthorQwen
Country🇨🇳 China
ReleasedAug 3, 2026
Overview
Cost rate
1x
Context
1M
Input
Output
Reasoning
Tool calling
Structured outputs
Pricing/M tokens
Input$2.00
Cached input$0.25
Output$6.00
Context
Context length1,000,000
Max output tokens131,072
Knowledge cutoff-
Performance
Throughput28 tok/s
Latency6.13s
Benchmarks
Head-to-head scores across reasoning, coding, and agentic indices, plus per-domain rankings.
Intelligence
62.5
Claude Opus 5 (Adaptive Reasoning, Xhigh Effort)
62.1
Claude Fable 5 (Adaptive Reasoning, Max Effort, Opus 4.8 Fallback)
59.5