Step 3.7 Flash
Stepfun🇨🇳
stepfun/step-3.7-flashStep 3.7 Flash is StepFun's multimodal vision-language model, built on Step 3.5 Flash with an added vision encoder for native image understanding. With a 256k context window, selectable reasoning levels, and high throughput, it targets agentic, coding, and search workflows that mix text and visual input.
コスト率
0.2x
コンテキスト
262K
リリース
2026年5月28日
入力
TextImageVideo
出力
Text
対応
推論ツール呼び出し構造化出力
得意なこと
このモデルが最も高くランクされているカテゴリーです。
上位カテゴリーが見つかりませんでした。
パフォーマンス
最近のリクエストで測定したレイテンシとスループットの中央値。
スループット150 tok/s
レイテンシ2.58s