Ling 3.1 Flash
InclusionAI🇨🇳
inclusionai/ling-3.1-flashLing 3.1 Flash is a hybrid reasoning mixture-of-experts model from inclusionAI, with 25B active parameters out of 560B total.
Cost rate
Free
Context
262K
Released
Oct 2, 2026
Input
Text
Output
Text
Support
ReasoningTool callingStructured outputs
Best at
The categories where this model ranks highest.
No top-ranked categories found.
Performance
Median latency and throughput measured across recent requests.
Throughput211 tok/s
Latency