StepFun
·Released on May 28, 2026

Step 3.7 Flash API Benchmarks, Pricing & Provider Data

Compare Step 3.7 Flash with another model

Choose a model to open its comparison page.

Share on X
LLM

Step 3.7 Flash benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.00005/request. Step 3.7 Flash free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, acti...

Quality
#75of 112
53.0
LMSpeed score
Speed
#31of 80
148char/s
22.50 s
Cost
#35of 186
$0.00005/ 1M · 8:1 in:out
$0.000008 in · $0.000042 out

Category Performance

Observed-capability estimates with uncertainty reported separately from independent benchmark families.

Coverage
6 / 8
Methodology
V3.0
Category PerformanceObserved-capability estimates with uncertainty reported separately from independent benchmark families.Agents53.3Coding48.7Reasoning51.8Knowledge42.1Math-Multilingual-Multimodal59Instruction following51.7
#1Multimodal59Provisional1/4 Measured dimensions
80% interval: 44.573.5
perception ocr59
simplevqa · simple_vqa / v_star · v_star
document spatialPrior only
visual reasoningPrior only
video actionPrior only
#2Agents53.3RatedGlobal rank #284/4 Measured dimensions
80% interval: 47.659.0
planning51.3
deep_planning · gert_labs
tool use58.6
tau · tau2_bench / toolathlon · toolathlon
environment execution51.2
browsecomp · browse_comp / deep_search_qa · deep_search_qa / gdpval_aa · benchlm_agentic_gdpval_aa / terminalbench · benchlm_agentic_terminal_bench2
recovery reliability52.1
apex_agents · apex_agents_aa / claw_eval · claw_eval
#3Reasoning51.8Estimated2/4 Measured dimensions
80% interval: 41.062.6
abstract logicPrior only
scientific causal52.7
critpt · critpt / gpqa · gpqa / hle · hle
multistep constraintsPrior only
evidence verification50.9
lcr · lcr
#4Instruction following51.7Provisional1/4 Measured dimensions
80% interval: 35.767.7
constraint followingPrior only
structured outputPrior only
novel instruction generalization51.7
ifbench · aa_if_bench
long multiturn instructionPrior only
#5Coding48.7Estimated3/4 Measured dimensions
80% interval: 39.458.0
code generation51.7
scicode · scicode
repository engineering48.2
swe_pro · swe_pro
debugging testingPrior only
tooling quality46.2
terminalbench · benchlm_coding_terminal_bench2
#6Knowledge42.1Provisional1/4 Measured dimensions
80% interval: 26.158.1
broad knowledge42.1
aa_omniscience · aa_omniscience_index
professional knowledgePrior only
factualityPrior only
retrieval open bookPrior only
No data:MathMultilingual

Specifications

Input and output token limits for this model, plus how it ranks on long-context understanding.

INPUT
262.1Ktokens
314.6 pages of text
OUTPUT
230.4Ktokens
8K128K1M4M
262.1K

Features

Technical Details

Input
Output
Released
May 2026
Tokenizer
Other
Architecture
text+image+video->text
Moderated
No
Supported parameters
frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p

Rankings

Excels at

It's decent at

Falls behind in

Detailed scores

Updated: Sep 13, 2026

Overall

undefined metric} other undefined metrics}}

Overall score53.0#75 / 112
Overall score
LMSpeed rank#75
Score53.0
UpdatedAug 20, 2026
Confidence1

Speed & latency

undefined metric} other undefined metrics}}

Output speed133.3 tok/s#31 / 80Time to first token1.58 s#44 / 80
Output speed
LMSpeed rank#31
Score133.3 tok/s
UpdatedSep 13, 2026
Confidence4
Time to first token
LMSpeed rank#44
Score1.58 s
UpdatedSep 13, 2026
Confidence4

Pricing

undefined metric} other undefined metrics}}

Input price$0.200/M#35 / 186Output price$1.15/M#49 / 186
Input price
LMSpeed rank#35
Score$0.200/M
UpdatedSep 13, 2026
Confidence4
Output price
LMSpeed rank#49
Score$1.15/M
UpdatedSep 13, 2026
Confidence4

Agents

V3.0

undefined metric} other undefined metrics}} · Rated

Score53.3#2880% interval47.6–59.04/4 Measured dimensionsAgentic score52.3#51 / 77Terminal-Bench 2.059.5#32 / 53
Dimensions and evidence
Planning & decomposition51.3
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
deep_planning · gert_labs · z 0.18 · q 1.00
Tool use58.6
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
tau · tau2_bench · z 1.32 · q 1.00
toolathlon · toolathlon · z 0.15 · q 1.00
Environment & long-horizon execution51.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
browsecomp · browse_comp · z -0.48 · q 1.00
deep_search_qa · deep_search_qa · z 0.85 · q 1.00
gdpval_aa · benchlm_agentic_gdpval_aa · z -0.37 · q 1.00
terminalbench · benchlm_agentic_terminal_bench2 · z -0.27 · q 1.00
Recovery & completion reliability52.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
apex_agents · apex_agents_aa · z -0.81 · q 1.00
claw_eval · claw_eval · z 1.27 · q 1.00
Agentic score
LMSpeed rank#51
Score52.3
UpdatedAug 20, 2026
Confidence1
Terminal-Bench 2.0V3 evidence
LMSpeed rank#32
Score59.5
UpdatedAug 20, 2026
Confidence1
BrowseCompV3 evidence
LMSpeed rank#21
Score75.8
UpdatedAug 20, 2026
Confidence1
DeepSearchQAV3 evidence
LMSpeed rank#4
Score92.8
UpdatedAug 20, 2026
Confidence1
GDPval-AA
LMSpeed rank#48
Score25.8
UpdatedAug 20, 2026
Confidence1
ToolathlonV3 evidence
LMSpeed rank#11
Score49.5
UpdatedAug 20, 2026
Confidence1
Claw-EvalV3 evidence
LMSpeed rank#6
Score67.1
UpdatedAug 20, 2026
Confidence1
HLE w/ tools
LMSpeed rank#8
Score47.2
UpdatedAug 20, 2026
Confidence1
Gert LabsV3 evidence
LMSpeed rank#20
Score51.6
UpdatedAug 20, 2026
Confidence1
AA Agentic Index
LMSpeed rank#43
Score21.7
UpdatedAug 20, 2026
Confidence1
Τ²-bench resultsV3 evidence
LMSpeed rank#3
Score98.5
UpdatedAug 20, 2026
Confidence1
GDPval-AAV3 evidence
LMSpeed rank#49
Score1018.0
UpdatedAug 20, 2026
Confidence1
APEX-Agents-AAV3 evidence
LMSpeed rank#17
Score14.8
UpdatedAug 20, 2026
Confidence1

Coding

V3.0

undefined metric} other undefined metrics}} · Estimated

Score48.780% interval39.4–58.03/4 Measured dimensionsSciCode43.9%#50 / 89Coding score41.2#74 / 87
Dimensions and evidence
Code generation51.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
scicode · scicode · z 0.06 · q 1.00
Repository engineering48.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
swe_pro · swe_pro · z -0.14 · q 1.00
Debugging & testingPrior only
Tool-assisted development & quality46.2
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
terminalbench · benchlm_coding_terminal_bench2 · z -0.49 · q 1.00
SciCodeV3 evidence
LMSpeed rank#50
Score43.9%
UpdatedSep 13, 2026
Confidence4
Coding score
LMSpeed rank#74
Score41.2
UpdatedAug 20, 2026
Confidence1
SWE-bench ProV3 evidence
LMSpeed rank#29
Score56.3
UpdatedAug 20, 2026
Confidence1
Terminal-Bench 2.0V3 evidence
LMSpeed rank#25
Score59.5
UpdatedAug 20, 2026
Confidence1
AA Coding Index
LMSpeed rank#57
Score39.6
UpdatedAug 20, 2026
Confidence1
AA-SciCodeV3 evidence
LMSpeed rank#75
Score40.0
UpdatedAug 20, 2026
Confidence1

Reasoning

V3.0

undefined metric} other undefined metrics}} · Estimated

Score51.880% interval41.0–62.62/4 Measured dimensionsGPQA80.9%#97 / 218HLE21.4%#83 / 216
Dimensions and evidence
Abstract logicPrior only
Scientific & causal reasoning52.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
critpt · critpt · z -0.12 · q 1.00
gpqa · gpqa · z 0.17 · q 1.00
hle · hle · z 0.29 · q 1.00
Multi-step constraintsPrior only
Evidence integration & verification50.9
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
lcr · lcr · z -0.23 · q 1.00
GPQAV3 evidence
LMSpeed rank#97
Score80.9%
UpdatedSep 13, 2026
Confidence4
HLEV3 evidence
LMSpeed rank#83
Score21.4%
UpdatedSep 13, 2026
Confidence4
AA-LCRV3 evidence
LMSpeed rank#65
Score69.7
UpdatedAug 20, 2026
Confidence1
CritPtV3 evidence
LMSpeed rank#60
Score2.3
UpdatedAug 20, 2026
Confidence1

Knowledge

V3.0

undefined metric} other undefined metrics}} · Provisional

Score42.180% interval26.1–58.11/4 Measured dimensionsArtificial Analysis Intelligence Index30.9#46 / 117AA-GPQA Diamond80.9#77 / 113
Dimensions and evidence
Broad knowledge42.1
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
aa_omniscience · aa_omniscience_index · z -1.14 · q 1.00
Professional knowledgePrior only
FactualityPrior only
Retrieval & open-book usePrior only
Artificial Analysis Intelligence IndexV3 evidence
LMSpeed rank#46
Score30.9
UpdatedAug 20, 2026
Confidence1
AA-GPQA Diamond
LMSpeed rank#77
Score80.9
UpdatedAug 20, 2026
Confidence1
AA-HLE
LMSpeed rank#69
Score21.4
UpdatedAug 20, 2026
Confidence1
AA-Omniscience Accuracy
LMSpeed rank#68
Score25.8
UpdatedAug 20, 2026
Confidence1
AA-Omniscience Hallucination Rate
LMSpeed rank#30
Score85.0
UpdatedAug 20, 2026
Confidence1

Math

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Foundational mathPrior only
Competition mathPrior only
Advanced proofsPrior only
Applied & tool-assisted mathPrior only

Multilingual

V3.0

undefined metric} other undefined metrics}} · No data

No data80% interval30.8–69.20/4 Measured dimensions
Dimensions and evidence
Cross-language understandingPrior only
Multilingual generationPrior only
Reasoning transferPrior only
Low-resource robustnessPrior only

Multimodal

V3.0

undefined metric} other undefined metrics}} · Provisional

Score5980% interval44.5–73.51/4 Measured dimensionsSimpleVQA79.2#2 / 8V*95.3#4 / 11
Dimensions and evidence
Perception & OCR59
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
simplevqa · simple_vqa · z 1.04 · q 1.00
v_star · v_star · z 0.52 · q 1.00
Document & spatial understandingPrior only
Visual reasoningPrior only
Video & grounded actionPrior only
SimpleVQAV3 evidence
LMSpeed rank#2
Score79.2
UpdatedAug 20, 2026
Confidence1
V*V3 evidence
LMSpeed rank#4
Score95.3
UpdatedAug 20, 2026
Confidence1
AA-MMMU-Pro
LMSpeed rank#32
Score75.3
UpdatedAug 20, 2026
Confidence1
Design Arena Website
LMSpeed rank#49
Score1202.0
UpdatedAug 20, 2026
Confidence1

Instruction following

V3.0

undefined metric} other undefined metrics}} · Provisional

Score51.780% interval35.7–67.71/4 Measured dimensionsAA-IFBench67.3#43 / 84
Dimensions and evidence
Constraint followingPrior only
Structured outputPrior only
Novel-instruction generalization51.7
undefined family} other undefined families}} · undefined metric} other undefined metrics}}
ifbench · aa_if_bench · z -0.07 · q 1.00
Multi-turn & long instructionsPrior only
AA-IFBenchV3 evidence
LMSpeed rank#43
Score67.3
UpdatedAug 20, 2026
Confidence1

OpenRouter endpoints

3 endpoints

Third-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.

Provider endpointInputOutput1d uptime30m latency30m throughputContext / output
DeepInfra
deepinfra
$0.160/M$0.920/M99.9%undefined tokens / undefined tokens
StepFun
stepfun/fp8
$0.200/M$1.15/M99.4%undefined tokens / undefined tokens
Novita
novita/fp8
$0.200/M$1.15/M99.4%undefined tokens / undefined tokens

Pricing Comparison

Compare Step 3.7 Flash API pricing across 48 providers. Prices range from $0.00005/request to $75.00/M. CM-API 公益站 offers the lowest rate at $0.00005/request. 9 providers offer free API credits or a free tier.

ProviderHealthModel VariantGroupInput ($/M)Output ($/M)Speed (t/s)First tokenAudit
L1
0%
step-3.7-flash
nvidia
-100%$0.0002/M
-100%$0.0011/M
71.7 t/s
17.60 s
L1
100%
L2
100%
step-3.7-flash
Archived
-8%$0.185/M
Cache read$0.037/M
-3%$1.11/M
36.5 t/s
2.76 s
L1
100%
step-3.7-flash
default
-26%$0.148/M
-23%$0.887/M
L1
100%
step-3.7-flash
default
-93%$0.014/M
-93%$0.079/M
L1
100%
step-3.7-flash
default
$7.50/M
$7.50/M
L1
100%
stepfun-ai/Step-3.7-Flash
default
$7.50/M
$7.50/M
L1
100%
L2
0%
step-3.7-flash
nvidia
$75.00/M
$75.00/M
L1
100%
Step-3.7-Flash
free
$0.00005/request
-
L1
100%
stepfun/step-3.7-flash:free
free
$0.00005/request
-
L1
100%
daipai/step-3.7-flash
按次福利模型
$0.0020/request
-
L1
99%
step-3.7-flash-free
default
Free
Free
L1
100%
stepfun-ai/Step-3.7-Flash
default
$0.479/request
-
L1
100%
step-3.7-flash
免费
-97%$0.0067/M
Cache read$0.0014/M
-96%$0.041/M
L1
100%
step-3.7-flash
default
$0.675/M
Cache read$0.135/M
$4.05/M
L1
99%
step-3.7-flash
default
Free
Free
L1
100%
stepfun-ai/step-3.7-flash
default
$75.00/M
$75.00/M
L1
99%
stepfun-ai/step-3.7-flash
0倍倍率分组
-93%$0.014/M
-96%$0.041/M
L1
0%
stepfun-ai/step-3.7-flash
2api
$3.00/M
$3.00/M
L1
61%
stepfun-ai/step-3.7-flash
default
$75.00/M
$75.00/M
L1
100%
step-3.7-flash
nvidia专区
-79%$0.041/M
Cache read$0.0082/M
-79%$0.236/M
Showing 20 model IDs of 31.

Alternatives & Similar Models

Frequently Asked Questions

What benchmark data does Step 3.7 Flash include?
LMSpeed shows Step 3.7 Flash benchmark context, API price, output speed, first-token latency, and provider data across 57 providers when those signals are available.
What is the Step 3.7 Flash API price?
Step 3.7 Flash has pricing from undefined provider} other undefined providers}}, ranging from $0.00005/request to $75.00/M. CM-API 公益站 has the lowest listed price.
What does the Step 3.7 Flash API pricing table include?
The Step 3.7 Flash API pricing table compares 57 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
Which provider has the cheapest Step 3.7 Flash API pricing?
CM-API 公益站 currently has the lowest listed Step 3.7 Flash price at $0.00005/request across undefined provider} other undefined providers}}.
Can I compare Step 3.7 Flash API price and speed together?
Yes. LMSpeed shows Step 3.7 Flash API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
Is Step 3.7 Flash API free?
Yes, Step 3.7 Flash free API options are available through 9 providerundefined other undefined} on LMSpeed, including Zero API, 初叶🍂Furry API, 兔子API, WSocket AI, WSocket AI. These providers offer free API credits or a free tier with no per-token charges.
Where can I get Step 3.7 Flash free API access?
LMSpeed currently lists 9 free API providerundefined other undefined} for Step 3.7 Flash: Zero API, 初叶🍂Furry API, 兔子API, WSocket AI, WSocket AI. Check each provider row before using it because free tier limits can change.

Also known as

Step-3.7-Flash[default]stepfun-ai/Step-3.7-Flashdaipai/step-3.7-flashlaohuang/stepfun-ai/step-3.7-flashstep-3-7-flash

Rankings are based on community-submitted tests and periodic health probes. Advisory only, not official data.Standard benchmark data may include BenchLM and other public sources.

Build with LMSpeed Data

Free

Free public API for LLM pricing, benchmarks & provider data

  • Real-time pricing & availability
  • Speed & latency benchmarks
  • 300+ models, 600+ providers
  • 1,000 requests/day
View API Documentation