Ling 3.0 Flash
InclusionAI🇨🇳
inclusionai/ling-3.0-flashLing 3.0 Flash is InclusionAI's (Ant Group) successor to Ling 2.6 Flash, a hybrid-reasoning Mixture-of-Experts model (~124B total, ~5.1B active) combining Kimi Delta Attention with Multi-Head Latent Attention for efficient long-range memory. It supports both thinking and non-thinking modes with a native 262K context, suited for fast general-purpose coding and agentic tasks.
Kostenfaktor
0.1x
Kontext
262K
Veröffentlicht
23. Juli 2026
Eingabe
Text
Ausgabe
Text
Unterstützung
SchlussfolgernTool-AufrufeStrukturierte Ausgaben
Worin es am besten ist
Die Kategorien, in denen dieses Modell am besten abschneidet.
Keine Top-Kategorien gefunden.
Leistung
Mittlere Latenz und Durchsatz, gemessen anhand aktueller Anfragen.
Durchsatz323 tok/s
Latenz