Step 3.7 Flash
Stepfun🇨🇳
stepfun/step-3.7-flashStep 3.7 Flash is StepFun's multimodal vision-language model, built on Step 3.5 Flash with an added vision encoder for native image understanding. With a 256k context window, selectable reasoning levels, and high throughput, it targets agentic, coding, and search workflows that mix text and visual input.
Taux de coût
0.2x
Contexte
262K
Date de sortie
28 mai 2026
Entrée
TextImageVideo
Sortie
Text
Prise en charge
RaisonnementAppel d'outilsSorties structurées
Meilleur en
Les catégories où ce modèle se classe le mieux.
Aucune catégorie classée en tête trouvée.
Performances
Latence et débit médians mesurés sur les requêtes récentes.
Débit177 tok/s
Latence