Step 3.7 Flash
Stepfun🇨🇳
stepfun/step-3.7-flashStep 3.7 Flash is StepFun's multimodal vision-language model, built on Step 3.5 Flash with an added vision encoder for native image understanding. With a 256k context window, selectable reasoning levels, and high throughput, it targets agentic, coding, and search workflows that mix text and visual input.
비용 비율
0.2x
컨텍스트
262K
출시일
2026년 5월 28일
입력
TextImageVideo
출력
Text
지원
추론도구 호출구조화된 출력
강점 분야
이 모델이 가장 높은 순위를 차지하는 카테고리입니다.
강점 카테고리를 찾을 수 없습니다.
성능
최근 요청 전반에서 측정한 중앙 지연 시간과 처리량입니다.
처리량150 tok/s
지연 시간2.58s