Mercury 2
Inception🇺🇸
inception/mercury-2Mercury 2 is Inception Labs' diffusion-based LLM that generates responses in parallel rather than token-by-token, delivering very high throughput (1000+ tokens/second) with a 128K context window. It offers tunable reasoning and native tool use, matching Claude Haiku/Gemini Flash-class quality at much faster speeds and lower cost, ideal for latency-sensitive agentic and coding tasks.
비용 비율
0.2x
컨텍스트
128K
출시일
2026년 3월 4일
입력
Text
출력
Text
지원
추론도구 호출구조화된 출력
강점 분야
이 모델이 가장 높은 순위를 차지하는 카테고리입니다.
소프트웨어 및 IT 서비스
#219 · 상위 53%
엔터테인먼트, 스포츠 및 미디어
#219 · 상위 53%
비즈니스, 경영 및 금융
#227 · 상위 56%
의료 및 헬스케어
#242 · 상위 63%
성능
최근 요청 전반에서 측정한 중앙 지연 시간과 처리량입니다.
처리량632 tok/s
지연 시간7.12s