Step 3.7 Flash
Stepfun🇨🇳
stepfun/step-3.7-flashStep 3.7 Flash is StepFun's multimodal vision-language model, built on Step 3.5 Flash with an added vision encoder for native image understanding. With a 256k context window, selectable reasoning levels, and high throughput, it targets agentic, coding, and search workflows that mix text and visual input.
Fascia di costo
0.2x
Contesto
262K
Rilasciato
28 mag 2026
Input
TextImageVideo
Output
Text
Supporto
RagionamentoChiamata di strumentiOutput strutturati
In cosa eccelle
Le categorie in cui questo modello ottiene i risultati migliori.
Nessuna categoria con classifica elevata trovata.
Prestazioni
Latenza e velocità mediane misurate sulle richieste recenti.
Velocità177 tok/s
Latenza