Step 3.7 Flash is StepFun's language model with a 262K context window and up to 256K output tokens, available from 4 providers, starting at $0.2 / 1M input and $1.15 / 1M output. A high-efficiency multimodal Mixture-of-Experts LLM combining a large language backbone with a vision encoder for native image and video understanding.
Specifications
Canonical IDstep-3-7-flash
TypeLanguage
StatusActive
CreatorStepFunStepFun
Providers
Context Window262K tokens
Max Output256K tokens
Input ModalitiesImageTextVideo
Output ModalitiesText
Reasoning Effortsdefault
Parameters201B
HuggingFace Likes370
HuggingFace Downloads (30d)50,187
HuggingFace Downloads (all-time)50,187
Release Date · 3 months ago
Benchmarks
Intelligence Index
#133
Coding Index
#91
GPQA
#140
HLE
#115
IFBench
#84
Time to First Token
#503
SciCode
#143
LCR
#100
TerminalBench Hard
#65
TAU2
#6
Output TPS
#130

Capabilities

Input3/5
Text
Image
Audio·
Video
PDF·
Output1/5
Text
Image·
Audio·
Video·
Embedding·
Capabilities5/13
Reasoning
Adaptive Reasoning·
Function Calling
Parallel Function Calling·
Structured Outputs
Native JSON Schema
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching
Assistant Prefill·

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandard
Input
$ / 1M
Output
$ / 1M
Cache Read
$ / 1M
DeepInfra logo
DeepInfra
deepinfra/stepfun-ai/Step-3.7-Flash
$0.2$1.15$0.04
Novita logo
Novita
novita/stepfun/step-3.7-flash
$0.2$1.15$0.04
OpenRouter logo
OpenRouter
stepfun/step-3.7-flash
$0.2$1.15$0.04
Vercel AI Gateway logo
Vercel AI Gateway
stepfun/step-3.7-flash
$0.2$1.15$0.04

Cost Calculator

US Dollar ($)
Preset:

Cheapest Instances to Run It

Cloud GPU instances that can host Step 3.7 Flash, ranked by cheapest on-demand price. The model needs about 483 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
p4de.24xlargeAWS8× A100640 GB$27.45/hr
Standard_ND96amsr_A100_v4Azure8× NVIDIA A100 (80GB)640 GB$32.77/hr
g7e.48xlargeAWS8× RTX PRO Server 6000768 GB$33.14/hr
7 more instances can run Step 3.7 Flash
Unlock the full ranked list and FP8 / INT4 quantization with a CloudPrice subscription.

Other Models

ModelTierReleasedContextInput / 1MOutput / 1M
Step 3 VL 10B
Step Image Edit 2
Step 1X Edit 1.2
Step 1X Edit

Model IDs

deepinfra/stepfun-ai/Step-3.7-Flash
novita/stepfun/step-3.7-flash
step-3-7-flash
stepfun-ai/Step-3.7-Flash
stepfun/step-3.7-flash