Step 3.7 Flash
Provider:StepFun
Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...
Key specs
- Context:262.1K
- Max output:230.4K
- Tokenizer:Other
- Released:2026-05-28
Token pricing
- Input price:$0.200 / 1M tokens
- Output price:$1.15 / 1M tokens
- Cache read:$0.040 / 1M tokens
- Blended price:$0.438 / 1M tokens
Modalities
Text, Image, Video
Use Cases
- Deep Reasoning & Analysis:Process long legal docs, academic papers, financial reports; multi-hop reasoning and evidence-chain tracing across huge contexts.(≈ $0.0092/call)
- Vision Understanding & Generation:Recognize data trends in charts, OCR scanned docs, analyze photos and produce structured reports.(≈ $0.0069/call)
- Chat & Support:FAQ auto-response, multi-turn dialogue, emotion detection — e-commerce, finance, gov.(≈ $0.0057/turn)
When to pick this model
Signals derived from public pricing and spec fields — not a benchmark. Validate with official docs and your own eval.
- Low input + output price: solid fit for high-frequency calls and batch processing.
- Long max output (>=30K tokens): strong fit for deep report / essay / document-drafting workflows.
- Multimodal (text+image): fit for image captioning, OCR, chart QA, vision-grounded reasoning.
- Video input support: fit for video understanding, frame QA, content moderation.
FAQ
- What is the token pricing of Step 3.7 Flash? Input price $0.200 / 1M tokens, output price $1.15 / 1M tokens, blended about $0.438 / 1M tokens. Refer to StepFun’s official page for the exact rate.
- What context window does Step 3.7 Flash support? Context window is 262.1K, max output about 230.4K, suitable for long documents, multi-turn chat and complex reasoning.
- What scenarios is Step 3.7 Flash best for? Based on its capability and pricing, Step 3.7 Flash fits best: Deep Reasoning & Analysis, Vision Understanding & Generation, Chat & Support. See the Use Cases section for details and cost estimates.
- What input/output modalities does Step 3.7 Flash support? Supports Text, Image, Video modality.
- Is Step 3.7 Flash a free model? Step 3.7 Flash is billed per token, not a free model.
- Which provider offers Step 3.7 Flash? Step 3.7 Flash is offered by StepFun.
More models from StepFun
Full info:Step 3.7 Flash