Ling 3.0 Flash
InclusionAI🇨🇳
inclusionai/ling-3.0-flashLing 3.0 Flash is InclusionAI's (Ant Group) successor to Ling 2.6 Flash, a hybrid-reasoning Mixture-of-Experts model (~124B total, ~5.1B active) combining Kimi Delta Attention with Multi-Head Latent Attention for efficient long-range memory. It supports both thinking and non-thinking modes with a native 262K context, suited for fast general-purpose coding and agentic tasks.
Cost rate
0.1x
Context
262K
Released
Jul 23, 2026
Input
Text
Output
Text
Support
ReasoningTool callingStructured outputs
Best at
The categories where this model ranks highest.
No top-ranked categories found.
Performance
Median latency and throughput measured across recent requests.
Throughput313 tok/s
Latency