Ling 3.0 Flash
InclusionAI🇨🇳
inclusionai/ling-3.0-flashLing 3.0 Flash is InclusionAI's (Ant Group) successor to Ling 2.6 Flash, a hybrid-reasoning Mixture-of-Experts model (~124B total, ~5.1B active) combining Kimi Delta Attention with Multi-Head Latent Attention for efficient long-range memory. It supports both thinking and non-thinking modes with a native 262K context, suited for fast general-purpose coding and agentic tasks.
费用比率
0.1x
上下文
262K
发布日期
2026年7月23日
输入
Text
输出
Text
支持
推理工具调用结构化输出
最擅长
该模型排名最高的类别。
未找到排名靠前的类别。
性能
根据近期请求测得的中位延迟和吞吐量。
吞吐量323 tok/s
延迟2.52s