Сравнить модели
Сравнивайте модели рядом по цене, контексту, возможностям и бенчмаркам.
Mercury 2.5 is Inception Labs' diffusion-based LLM (dLLM), generating tokens in parallel rather than sequentially to reach about 1,107 tokens/second with a 260K-token context window, and is billed as the largest diffusion language model trained to date. It delivers a roughly 40% intelligence gain over Mercury 2 (comparable to cost-optimized frontier models like Claude Haiku 4.5, Gemini 3.5 Flash-Lite, and GPT-5.6 Luna Low) with tunable reasoning, parallel tool calls, and schema-aligned JSON, suiting latency-sensitive search agents, voice pipelines, and coding subagents.
Claude Haiku 4.5 is Anthropic's fast, affordable small model in the Claude 4 generation, balancing strong instruction-following and coding ability with low latency. It is designed for high-throughput agentic and chat applications where speed and cost efficiency are priorities.
Бенчмарки
Сравнительные показатели по индексам рассуждений, программирования и агентности, а также рейтинги по отдельным доменам.