Step 3.7 Flash
Stepfun🇨🇳
stepfun/step-3.7-flashStep 3.7 Flash is StepFun's multimodal vision-language model, built on Step 3.5 Flash with an added vision encoder for native image understanding. With a 256k context window, selectable reasoning levels, and high throughput, it targets agentic, coding, and search workflows that mix text and visual input.
Taxa de custo
0.2x
Contexto
262K
Lançamento
28 de mai. de 2026
Entrada
TextImageVideo
Saída
Text
Suporte
RaciocínioChamada de ferramentasSaídas estruturadas
Melhor em
As categorias em que este modelo tem o melhor desempenho.
Nenhuma categoria de destaque encontrada.
Desempenho
Latência e vazão medianas medidas em solicitações recentes.
Vazão177 tok/s
Latência