Confronta modelli
Metti i modelli a confronto per prezzo, contesto, capacità e benchmark.
Mercury 2.5 is Inception Labs' diffusion-based LLM (dLLM), generating tokens in parallel rather than sequentially to reach about 1,107 tokens/second with a 260K-token context window, and is billed as the largest diffusion language model trained to date. It delivers a roughly 40% intelligence gain over Mercury 2 (comparable to cost-optimized frontier models like Claude Haiku 4.5, Gemini 3.5 Flash-Lite, and GPT-5.6 Luna Low) with tunable reasoning, parallel tool calls, and schema-aligned JSON, suiting latency-sensitive search agents, voice pipelines, and coding subagents.
Claude Haiku 4.5 is Anthropic's fast, affordable small model in the Claude 4 generation, balancing strong instruction-following and coding ability with low latency. It is designed for high-throughput agentic and chat applications where speed and cost efficiency are priorities.
Benchmark
Punteggi testa a testa negli indici di ragionamento, programmazione e capacità agentiche, oltre alle classifiche per dominio.