endueendue

StepFun

Step 3.7 Flash

stepfun/step-3.7-flash

Efficient MoE with native image and video understanding.

MultimodalText, Bilder, Video → TextSofort in endue nutzbar
  • Audio & Video
  • Geringe Kosten
  • Bildverständnis

Überblick

Geeignet für

Verfasst von endue

  • Audio & Video
  • Geringe Kosten
  • Bildverständnis

Weniger geeignet für

Aus den Spezifikationen des Modells

Nichts in den Spezifikationen spricht gegen übliche Agentenaufgaben.

Fähigkeiten

Tool useUnterstützt
Structured outputUnterstützt
JSON modeUnterstützt
ReasoningUnterstützt
VisionUnterstützt
Audio inputNicht unterstützt
Video inputUnterstützt
File inputNicht unterstützt
Long contextUnterstützt

Spezifikationen

Modell-ID
stepfun/step-3.7-flash
Modellfamilie
Step
Kontextfenster
262K Tokens
Eingabe
Text, Bilder, Video
Ausgabe
Text
Veröffentlicht
2026-05-28