Mercury 2
Inception🇺🇸
inception/mercury-2Mercury 2 is Inception Labs' diffusion-based LLM that generates responses in parallel rather than token-by-token, delivering very high throughput (1000+ tokens/second) with a 128K context window. It offers tunable reasoning and native tool use, matching Claude Haiku/Gemini Flash-class quality at much faster speeds and lower cost, ideal for latency-sensitive agentic and coding tasks.
Taux de coût
0.2x
Contexte
128K
Date de sortie
4 mars 2026
Entrée
Text
Sortie
Text
Prise en charge
RaisonnementAppel d'outilsSorties structurées
Meilleur en
Les catégories où ce modèle se classe le mieux.
Logiciels et services informatiques
#219 · top 53 %
Divertissement, sport et médias
#219 · top 53 %
Affaires, gestion et finance
#227 · top 56 %
Médecine et santé
#242 · top 63 %
Performances
Latence et débit médians mesurés sur les requêtes récentes.
Débit697 tok/s
Latence