Ling 3.0 Flash
InclusionAI🇨🇳
inclusionai/ling-3.0-flashLing 3.0 Flash is InclusionAI's (Ant Group) successor to Ling 2.6 Flash, a hybrid-reasoning Mixture-of-Experts model (~124B total, ~5.1B active) combining Kimi Delta Attention with Multi-Head Latent Attention for efficient long-range memory. It supports both thinking and non-thinking modes with a native 262K context, suited for fast general-purpose coding and agentic tasks.
비용 비율
0.1x
컨텍스트
262K
출시일
2026년 7월 23일
입력
Text
출력
Text
지원
추론도구 호출구조화된 출력
강점 분야
이 모델이 가장 높은 순위를 차지하는 카테고리입니다.
강점 카테고리를 찾을 수 없습니다.
성능
최근 요청 전반에서 측정한 중앙 지연 시간과 처리량입니다.
처리량323 tok/s
지연 시간2.52s