Step 3.7 Flash API Benchmarks, Pricing & Provider Data
Compare Step 3.7 Flash with another model
Choose a model to open its comparison page.
Step 3.7 Flash benchmark, API pricing, and provider data cover undefined API provider} other undefined API providers}}, with prices starting at $0.00005/request. Step 3.7 Flash free API options are available from undefined provider} other undefined providers}}. The page also shows measured API speed and first-token latency.
Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, acti...
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 6 / 8
- Methodology
- V3.0
#1Multimodal59Provisional1/4 Measured dimensions
#2Agents53.3RatedGlobal rank #284/4 Measured dimensions
#3Reasoning51.8Estimated2/4 Measured dimensions
#4Instruction following51.7Provisional1/4 Measured dimensions
#5Coding48.7Estimated3/4 Measured dimensions
#6Knowledge42.1Provisional1/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- May 2026
- Tokenizer
- Other
- Architecture
- text+image+video->text
- Moderated
- No
- Supported parameters
- frequency_penaltyinclude_reasoninglogit_biaslogprobsmax_tokensmin_ppresence_penaltyreasoningreasoning_effortrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Rankings
Excels at
It's decent at
Falls behind in
Detailed scores
Updated: Sep 13, 2026Overall
undefined metric} other undefined metrics}}
Overall score53.0#75 / 112
Overall
undefined metric} other undefined metrics}}
Speed & latency
undefined metric} other undefined metrics}}
Output speed133.3 tok/s#31 / 80Time to first token1.58 s#44 / 80
Speed & latency
undefined metric} other undefined metrics}}
Pricing
undefined metric} other undefined metrics}}
Input price$0.200/M#35 / 186Output price$1.15/M#49 / 186
Pricing
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Score53.3#2880% interval47.6–59.04/4 Measured dimensionsAgentic score52.3#51 / 77Terminal-Bench 2.059.5#32 / 53
Agents
V3.0undefined metric} other undefined metrics}} · Rated
Coding
V3.0undefined metric} other undefined metrics}} · Estimated
Score48.780% interval39.4–58.03/4 Measured dimensionsSciCode43.9%#50 / 89Coding score41.2#74 / 87
Coding
V3.0undefined metric} other undefined metrics}} · Estimated
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Score51.880% interval41.0–62.62/4 Measured dimensionsGPQA80.9%#97 / 218HLE21.4%#83 / 216
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Score42.180% interval26.1–58.11/4 Measured dimensionsArtificial Analysis Intelligence Index30.9#46 / 117AA-GPQA Diamond80.9#77 / 113
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Math
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Math
V3.0undefined metric} other undefined metrics}} · No data
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · Provisional
Score5980% interval44.5–73.51/4 Measured dimensionsSimpleVQA79.2#2 / 8V*95.3#4 / 11
Multimodal
V3.0undefined metric} other undefined metrics}} · Provisional
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
Score51.780% interval35.7–67.71/4 Measured dimensionsAA-IFBench67.3#43 / 84
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
OpenRouter endpoints
3 endpointsThird-party OpenRouter endpoint data, shown separately from LMSpeed measurements. Some 30-minute live performance fields only appear after syncing with an OpenRouter API key.
| Provider endpoint | Input | Output | 1d uptime | 30m latency | 30m throughput | Context / output |
|---|---|---|---|---|---|---|
DeepInfra deepinfra | $0.160/M | $0.920/M | 99.9% | — | — | undefined tokens / undefined tokens |
StepFun stepfun/fp8 | $0.200/M | $1.15/M | 99.4% | — | — | undefined tokens / undefined tokens |
Novita novita/fp8 | $0.200/M | $1.15/M | 99.4% | — | — | undefined tokens / undefined tokens |
Pricing Comparison
Compare Step 3.7 Flash API pricing across 48 providers. Prices range from $0.00005/request to $75.00/M. CM-API 公益站 offers the lowest rate at $0.00005/request. 9 providers offer free API credits or a free tier.
| Provider | Health | Model Variant | Group | Input ($/M) | Output ($/M) | Speed (t/s) | First token | Audit |
|---|---|---|---|---|---|---|---|---|
L1 0% | step-3.7-flash | nvidia | -100%$0.0002/M | -100%$0.0011/M | 71.7 t/s | 17.60 s | — | |
L1 100% L2 100% | step-3.7-flash | Archived | -8%$0.185/M Cache read$0.037/M | -3%$1.11/M | 36.5 t/s | 2.76 s | — | |
L1 100% | step-3.7-flash | default | -26%$0.148/M | -23%$0.887/M | — | — | — | |
L1 100% | step-3.7-flash | default | -93%$0.014/M | -93%$0.079/M | — | — | — | |
L1 100% | step-3.7-flash | default | $7.50/M | $7.50/M | — | — | — | |
L1 100% | stepfun-ai/Step-3.7-Flash | default | $7.50/M | $7.50/M | — | — | — | |
L1 100% L2 0% | step-3.7-flash | nvidia | $75.00/M | $75.00/M | — | — | — | |
L1 100% | Step-3.7-Flash | free | $0.00005/request | - | — | — | — | |
L1 100% | stepfun/step-3.7-flash:free | free | $0.00005/request | - | — | — | — | |
L1 100% | daipai/step-3.7-flash | 按次福利模型 | $0.0020/request | - | — | — | — | |
兔子API Free | L1 99% | step-3.7-flash-free | default | Free | Free | — | — | — |
L1 100% | stepfun-ai/Step-3.7-Flash | default | $0.479/request | - | — | — | — | |
L1 100% | step-3.7-flash | 免费 | -97%$0.0067/M Cache read$0.0014/M | -96%$0.041/M | — | — | — | |
L1 100% | step-3.7-flash | default | $0.675/M Cache read$0.135/M | $4.05/M | — | — | — | |
Moyanjdc API Free | L1 99% | step-3.7-flash | default | Free | Free | — | — | — |
L1 100% | stepfun-ai/step-3.7-flash | default | $75.00/M | $75.00/M | — | — | — | |
L1 99% | stepfun-ai/step-3.7-flash | 0倍倍率分组 | -93%$0.014/M | -96%$0.041/M | — | — | — | |
L1 0% | stepfun-ai/step-3.7-flash | 2api | $3.00/M | $3.00/M | — | — | — | |
L1 61% | stepfun-ai/step-3.7-flash | default | $75.00/M | $75.00/M | — | — | — | |
L1 100% | step-3.7-flash | nvidia专区 | -79%$0.041/M Cache read$0.0082/M | -79%$0.236/M | — | — | — |
Alternatives & Similar Models
GLM-5.1
glm-5-1
Zhipu GLM-5.1 is a next-generation GLM model aimed at frontier reasoning, coding, and bilingual agent applications.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
MiniMax M2.7
minimax-m2-7
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
DeepSeek V4 Pro
deepseek-v4-pro
DeepSeek V4 Pro is the professional-tier DeepSeek V4 model, targeting frontier reasoning, coding, and agent workflows with maximum capability.
MiniMax M3
minimax-m3
MiniMax M3 is MiniMax next-generation large language model, designed for advanced reasoning, long-context understanding, and high-quality multilingual dialogue.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
Frequently Asked Questions
- What benchmark data does Step 3.7 Flash include?
- LMSpeed shows Step 3.7 Flash benchmark context, API price, output speed, first-token latency, and provider data across 57 providers when those signals are available.
- What is the Step 3.7 Flash API price?
- Step 3.7 Flash has pricing from undefined provider} other undefined providers}}, ranging from $0.00005/request to $75.00/M. CM-API 公益站 has the lowest listed price.
- What does the Step 3.7 Flash API pricing table include?
- The Step 3.7 Flash API pricing table compares 57 providers by input price, output price, free tier status, speed, first-token latency, and recent health data when available.
- Which provider has the cheapest Step 3.7 Flash API pricing?
- CM-API 公益站 currently has the lowest listed Step 3.7 Flash price at $0.00005/request across undefined provider} other undefined providers}}.
- Can I compare Step 3.7 Flash API price and speed together?
- Yes. LMSpeed shows Step 3.7 Flash API price, output speed, first-token latency, and provider health on the same page so you can compare cost and performance together.
- Is Step 3.7 Flash API free?
- Yes, Step 3.7 Flash free API options are available through 9 providerundefined other undefined} on LMSpeed, including Zero API, 初叶🍂Furry API, 兔子API, WSocket AI, WSocket AI. These providers offer free API credits or a free tier with no per-token charges.
- Where can I get Step 3.7 Flash free API access?
- LMSpeed currently lists 9 free API providerundefined other undefined} for Step 3.7 Flash: Zero API, 初叶🍂Furry API, 兔子API, WSocket AI, WSocket AI. Check each provider row before using it because free tier limits can change.
