StepFun
·Released on May 28, 2026

Step 3.7 Flash API Benchmarks, Pricing & Provider Data

Compare Step 3.7 Flash with another model

Choose a model to open its comparison page.

Share on X
LLM

Step 3.7 Flash benchmark, API pricing, and provider data cover 69 API providers, with prices starting at $0.0010/request. Step 3.7 Flash free API options are available from 10 providers. The page also shows measured API speed and first-token latency.

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, acti...

Quality
#63of 102
53.0
LMSpeed score
Speed
#42of 79
148char/s
22.50 s
Cost
#34of 183
$0.0010/ 1M · 8:1 in:out
$0.0002 in · $0.0008 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
6 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents53.6Coding49Reasoning52.7Knowledge42Math-Multilingual-Multimodal59Instruction following51.6
#1Multimodal59Provisional1/4 Measured dimensions
80% interval: 44.573.5
perception ocr59
simplevqa · simple_vqa / v_star · v_star
document spatialPrior only
visual reasoningPrior only
video actionPrior only
#2Agents53.6RatedGlobal rank #244/4 Measured dimensions
80% interval: 47.959.3
planning51.3
deep_planning · gert_labs
tool use58.7
tau · tau2_bench / toolathlon · toolathlon
environment execution51.8
browsecomp · browse_comp / deep_search_qa · deep_search_qa / gdpval_aa · benchlm_agentic_gdpval_aa / terminalbench · benchlm_agentic_terminal_bench2
recovery reliability52.5
apex_agents · apex_agents_aa / claw_eval · claw_eval
#3Reasoning52.7Estimated2/4 Measured dimensions
80% interval: 41.963.5
abstract logicPrior only
scientific causal53
critpt · critpt / gpqa · gpqa / hle · hle
multistep constraintsPrior only
evidence verification52.4
lcr · lcr
#4Instruction following51.6Provisional1/4 Measured dimensions
80% interval: 35.667.6
constraint followingPrior only
structured outputPrior only
novel instruction generalization51.6
ifbench · aa_if_bench
long multiturn instructionPrior only
#5Coding49Estimated3/4 Measured dimensions
80% interval: 39.758.3
code generation52.2
scicode · scicode
repository engineering48.5
swe_pro · swe_pro
debugging testingPrior only
tooling quality46.3
terminalbench · benchlm_coding_terminal_bench2
#6Knowledge42Provisional1/4 Measured dimensions
80% interval: 26.058.0
broad knowledge42
aa_omniscience · aa_omniscience_index
professional knowledgePrior only
factualityPrior only
retrieval open bookPrior only
No data:MathMultilingual

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
262.1Ktokens
314.6 pages of text
OUTPUT
230.4Ktokens
8K128K1M4M
262.1K

Features

Technical Details

Input
Output
Released
May 2026
Tokenizer
Other
Architecture
text+image+video->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Sep 3, 2026

Overall

undefined metric} other undefined metrics}}

Overall score53.0#63 / 102
Overall score
LMSpeed rank#63
Score53.0
UpdatedAug 20, 2026
Confidence1

Speed & latency

undefined metric} other undefined metrics}}

Output speed94.7 tok/s#42 / 79Time to first token1.58 s#49 / 79
Output speed
LMSpeed rank#42
Score94.7 tok/s
UpdatedSep 3, 2026
Confidence4
Time to first token
LMSpeed rank#49
Score1.58 s
UpdatedSep 3, 2026
Confidence4

Pricing

undefined metric} other undefined metrics}}

Input price$0.200/M#34 / 183Output price$1.15/M#49 / 183
Input price
LMSpeed rank#34
Score$0.200/M
UpdatedSep 3, 2026
Confidence4
Output price
LMSpeed rank#49
Score$1.15/M
UpdatedSep 3, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score53.6#2480% interval47.9–59.34/4 Measured dimensionsAgentic score52.3#47 / 71Terminal-Bench 2.059.5#32 / 53
Dimensions and evidence
Planning & decomposition51.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
deep_planning · gert_labs · z 0.18 · q 1.00
Tool use58.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
tau · tau2_bench · z 1.29 · q 1.00
toolathlon · toolathlon · z 0.15 · q 1.00
Environment & long-horizon execution51.8
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
browsecomp · browse_comp · z -0.37 · q 1.00
deep_search_qa · deep_search_qa · z 0.94 · q 1.00
gdpval_aa · benchlm_agentic_gdpval_aa · z -0.41 · q 1.00
terminalbench · benchlm_agentic_terminal_bench2 · z -0.27 · q 1.00
Recovery & completion reliability52.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
apex_agents · apex_agents_aa · z -0.81 · q 1.00
claw_eval · claw_eval · z 1.27 · q 1.00
Agentic score
LMSpeed rank#47
Score52.3
UpdatedAug 20, 2026
Confidence1
Terminal-Bench 2.0V3 evidence
LMSpeed rank#32
Score59.5
UpdatedAug 20, 2026
Confidence1
BrowseCompV3 evidence
LMSpeed rank#20
Score75.8
UpdatedAug 20, 2026
Confidence1
DeepSearchQAV3 evidence
LMSpeed rank#4
Score92.8
UpdatedAug 20, 2026
Confidence1
GDPval-AA
LMSpeed rank#45
Score25.8
UpdatedAug 20, 2026
Confidence1
ToolathlonV3 evidence
LMSpeed rank#11
Score49.5
UpdatedAug 20, 2026
Confidence1
Claw-EvalV3 evidence
LMSpeed rank#6
Score67.1
UpdatedAug 20, 2026
Confidence1
HLE w/ tools
LMSpeed rank#6
Score47.2
UpdatedAug 20, 2026
Confidence1
Gert LabsV3 evidence
LMSpeed rank#20
Score51.6
UpdatedAug 20, 2026
Confidence1
AA Agentic Index
LMSpeed rank#46
Score21.7
UpdatedAug 20, 2026
Confidence1
Τ²-bench resultsV3 evidence
LMSpeed rank#3
Score98.5
UpdatedAug 20, 2026
Confidence1
GDPval-AAV3 evidence
LMSpeed rank#45
Score1018.0
UpdatedAug 20, 2026
Confidence1
APEX-Agents-AAV3 evidence
LMSpeed rank#17
Score14.8
UpdatedAug 20, 2026
Confidence1

Coding

V3.0

undefined metric} other undefined metrics}} · Estimated

Score4980% interval39.7–58.33/4 Measured dimensionsSciCode40.0%#87 / 208Coding score41.2#72 / 82
Dimensions and evidence
Code generation52.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
scicode · scicode · z 0.17 · q 1.00
Repository engineering48.5
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_pro · swe_pro · z -0.12 · q 1.00
Debugging & testingPrior only
Tool-assisted development & quality46.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
terminalbench · benchlm_coding_terminal_bench2 · z -0.49 · q 1.00
SciCodeV3 evidence
LMSpeed rank#87
Score40.0%
UpdatedSep 3, 2026
Confidence4
Coding score
LMSpeed rank#72
Score41.2
UpdatedAug 20, 2026
Confidence1
SWE-bench ProV3 evidence
LMSpeed rank#29
Score56.3
UpdatedAug 20, 2026
Confidence1
Terminal-Bench 2.0V3 evidence
LMSpeed rank#25
Score59.5
UpdatedAug 20, 2026
Confidence1
AA Coding Index
LMSpeed rank#53
Score39.6
UpdatedAug 20, 2026
Confidence1
AA-SciCodeV3 evidence
LMSpeed rank#70
Score40.0
UpdatedAug 20, 2026
Confidence1

Reasoning

V3.0

undefined metric} other undefined metrics}} · Estimated

Score52.780% interval41.9–63.52/4 Measured dimensionsGPQA80.9%#94 / 215HLE21.4%#79 / 212
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning53
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z -0.09 · q 1.00
gpqa · gpqa · z 0.21 · q 1.00
hle · hle · z 0.34 · q 1.00
Multi-step constraintsPrior only
Evidence integration & verification52.4
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
lcr · lcr · z -0.04 · q 1.00
GPQAV3 evidence
LMSpeed rank#94
Score80.9%
UpdatedSep 3, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#79
Score21.4%
UpdatedSep 3, 2026
Confidence4
AA-LCRV3 evidence
LMSpeed rank#53
Score69.7
UpdatedAug 20, 2026
Confidence1
CritPtV3 evidence
LMSpeed rank#56
Score2.3
UpdatedAug 20, 2026
Confidence1

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score4280% interval26.0–58.01/4 Measured dimensionsArtificial Analysis Intelligence Index30.9#71 / 112AA-GPQA Diamond80.9#73 / 109
Dimensions and evidence
Broad knowledge42
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
aa_omniscience · aa_omniscience_index · z -1.14 · q 1.00
Professional knowledgePrior only
FactualityPrior only
Retrieval & open-book usePrior only
Artificial Analysis Intelligence IndexV3 evidence
LMSpeed rank#71
Score30.9
UpdatedAug 20, 2026
Confidence1
AA-GPQA Diamond
LMSpeed rank#73
Score80.9
UpdatedAug 20, 2026
Confidence1
AA-HLE
LMSpeed rank#65
Score21.4
UpdatedAug 20, 2026
Confidence1
AA-Omniscience Accuracy
LMSpeed rank#66
Score25.8
UpdatedAug 20, 2026
Confidence1
AA-Omniscience Hallucination Rate
LMSpeed rank#31
Score85.0
UpdatedAug 20, 2026
Confidence1

Math

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Foundational mathPrior only
Competition mathPrior only
Advanced proofsPrior only
Applied & tool-assisted mathPrior only

Multilingual

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Cross-language understandingPrior only
Multilingual generationPrior only
Reasoning transferPrior only
Low-resource robustnessPrior only

Multimodal

V3.0

undefined metric} other undefined metrics}} · Provisional

Score5980% interval44.5–73.51/4 Measured dimensionsSimpleVQA79.2#2 / 8V*95.3#4 / 11
Dimensions and evidence
Perception & OCR59
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
simplevqa · simple_vqa · z 1.04 · q 1.00
v_star · v_star · z 0.52 · q 1.00
Document & spatial understandingPrior only
Visual reasoningPrior only
Video & grounded actionPrior only
SimpleVQAV3 evidence
LMSpeed rank#2
Score79.2
UpdatedAug 20, 2026
Confidence1
V*V3 evidence
LMSpeed rank#4
Score95.3
UpdatedAug 20, 2026
Confidence1
AA-MMMU-Pro
LMSpeed rank#30
Score75.3
UpdatedAug 20, 2026
Confidence1
Design Arena Website
LMSpeed rank#47
Score1202.0
UpdatedAug 20, 2026
Confidence1

Instruction following

V3.0

undefined metric} other undefined metrics}} · Provisional

Score51.680% interval35.6–67.61/4 Measured dimensionsAA-IFBench67.3#44 / 84
Dimensions and evidence
Constraint followingPrior only
Structured outputPrior only
Novel-instruction generalization51.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifbench · aa_if_bench · z -0.07 · q 1.00
Multi-turn & long instructionsPrior only
AA-IFBenchV3 evidence
LMSpeed rank#44
Score67.3
UpdatedAug 20, 2026
Confidence1

OpenRouter endpoints

3 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
DeepInfra
deepinfra
$0.160/M$0.920/M99.9%undefined tokens / undefined tokens
StepFun
stepfun/fp8
$0.200/M$1.15/M99.1%undefined tokens / undefined tokens
Novita
novita/fp8
$0.200/M$1.15/M97.7%undefined tokens / undefined tokens

Pricing Comparison

Compare Step 3.7 Flash API pricing across 59 providers. Prices range from $0.0010/request to $1998.00/M. CM-API 公益站 offers the lowest rate at $0.0010/request. 10 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
100%
step-3.7-flash-free
default
Free
Free
L1
100%
stepfun/step-3.7-flash:free
free
$0.0010/request
-
L1
100%
Step-3.7-Flash
国产模型
$0.010/request
-
L1
100%
step-3.7-flash
default
-26%$0.148/M
-23%$0.887/M
L1
100%
step-3.7-flash
default
$1.40/M
$8.05/M
L1
100%
L2
55%
step-3.7-flash
nvidia
$75.00/M
$75.00/M
L1
99%
stepfun-ai/Step-3.7-Flash
default
$0.0048/request
-
L1
100%
step-3.7-flash
nvidia
$0.400/M
Cache read$0.080/M
$2.30/M
L1
100%
L2
81%
step-3.7-flash
default
$75.00/M
$75.00/M
L1
100%
stepfun-ai/step-3.7-flash
default
$375.00/M
$375.00/M
L1
100%
step-3.7-flash
default
$375.00/M
$375.00/M
L1
100%
stepfun/step-3.7-flash-free
default
$375.00/M
$375.00/M
L1
99%
step-3.7-flash
default
-93%$0.014/M
-93%$0.079/M
L1
99%
step-3.7-flash
default
Free
Free
L1
99%
daipai/step-3.7-flash
按次福利模型
$0.0020/request
-
L1
100%
stepfun-ai/step-3.7-flash
default
$75.00/M
$75.00/M
L1
99%
step-3.7-flash
default
$0.661/M
Cache read$0.132/M
$3.97/M
L1
99%
step-3.7-flash
91vip
$1.35/M
$7.77/M
L1
99%
stepfun-ai/step-3.7-flash
default
$75.00/M
$75.00/M
L1
99%
step-3.7-flash
免费
$0.375/M
-67%$0.375/M
Showing 20 model IDs of 37.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Step 3.7 Flash include?
LMSpeed shows Step 3.7 Flash benchmark context, API price, output speed, first-token latency, and provider data across 69 providers when those signals are available.
What is the Step 3.7 Flash API price?
Step 3.7 Flash has pricing from 69 providers, ranging from $0.0010/request to $1998.00/M. CM-API 公益站 has the lowest listed price.
What does the Step 3.7 Flash API pricing table include?
The Step 3.7 Flash API pricing table compares 69 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Step 3.7 Flash API pricing?
CM-API 公益站 currently has the lowest listed Step 3.7 Flash price at $0.0010/request across 69 providers.
Can I compare Step 3.7 Flash API price and speed together?
Yes. LMSpeed shows Step 3.7 Flash API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is Step 3.7 Flash API free?
Yes, Step 3.7 Flash free API options are available through 10 providerundefined other undefined} on LMSpeed, including WSocket AI, Dext API, 兔子API, 兔子API, Zero API. These providers offer free API credits or a free tier with no per-token charges.
Where can I get Step 3.7 Flash free API access?
LMSpeed currently lists 10 free API providerundefined other undefined} for Step 3.7 Flash: WSocket AI, Dext API, 兔子API, 兔子API, Zero API. Check each provider row before using it because free tier limits can change.

Also known as

Step-3.7-Flashdaipai/step-3.7-flashstep-3-7-flashstep-3.7-flashstep-3.7-flash-free

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation