Mercury 2
Inception🇺🇸
inception/mercury-2Mercury 2 is Inception Labs' diffusion-based LLM that generates responses in parallel rather than token-by-token, delivering very high throughput (1000+ tokens/second) with a 128K context window. It offers tunable reasoning and native tool use, matching Claude Haiku/Gemini Flash-class quality at much faster speeds and lower cost, ideal for latency-sensitive agentic and coding tasks.
コスト率
0.2x
コンテキスト
128K
リリース
2026年3月4日
入力
Text
出力
Text
対応
推論ツール呼び出し構造化出力
得意なこと
このモデルが最も高くランクされているカテゴリーです。
ソフトウェア・ITサービス
#219 · 上位53%
エンターテインメント・スポーツ・メディア
#219 · 上位53%
ビジネス・経営・金融
#227 · 上位56%
医療・ヘルスケア
#242 · 上位63%
パフォーマンス
最近のリクエストで測定したレイテンシとスループットの中央値。
スループット632 tok/s
レイテンシ7.12s