Step 5 Preview is StepFun's 600B sparse MoE frontier model for production-scale agents, activating 27B parameters per token. It is built for software engineering, long-horizon tool use, research, professional knowledge work, and finance, with native text, image, and video understanding and a 1M-token context window. ⚠️ Note: This model routes through StepFun, so privacy and logging guarantees may be limited.
Added Sep 20, 2026
Context Window
1.0M
Max Output
1.0M
Input Price (Auto)
$1.00/1M
Output Price (Auto)
$2.70/1M
Cache Read (Auto)
$0.050/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
43.7
Agentic work
AutomationBench-AA
Workflow automation with guardrail penalties
51.0%
Better than 74% of models compared
AutomationBench-AA Tasks Completed
Fully completed workflows without guardrail violations
17.7%
Better than 29% of models compared
AA-Briefcase
Agentic knowledge work (Elo)
1433 Elo
Better than 83% of models compared
GDPval-AA v2
Economically valuable tasks (Elo)
1566 Elo
Better than 91% of models compared
Document reasoning
GDP.pdf
Professional PDF reasoning: all-pass rate
14.8%
Better than 61% of models compared
AA-LCR v1.1
Long context reasoning with updated grading
88.3%
Better than 99% of models compared
Reasoning
HLE
Humanity's Last Exam
46.5%
Better than 96% of models compared
Coding
Terminal-Bench v4.0
Practical coding and terminal tasks
33.3%
Better than 88% of models compared
SciCode
Python programming for scientific computing
58.9%
Better than 96% of models compared
Legacy benchmarks
AA-LCR (unversioned / legacy)
Long context reasoning evaluation
88.3%
Better than 99% of models compared
Last updated Sep 20, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Step 5 Preview with similar models from the same provider or model family.
Step 3.7 Flash Thinking
stepfun/step-3.7-flash:thinkingStep 3.7 Flash Thinking is StepFun's high-efficiency multimodal MoE model with visible reasoning enabled for deeper agentic coding, long-context reasoning, tool use, and native image/video understanding. ⚠️ Note: This model routes through StepFun, so privacy and logging guarantees may be limited.
Step 3.5 Flash 2603
stepfun-ai/step-3.5-flash-2603Step 3.5 Flash 2603 is optimized for high-frequency agentic and coding workflows with improved token efficiency and faster reasoning. NOTE: This model runs via StepFun, which may log and train on your prompts.
Step 3.5 Flash
stepfun-ai/step-3.5-flashStepFun's most capable open-source reasoning model with visible reasoning traces. Built on a sparse Mixture-of-Experts architecture with 196B total parameters and only 11B active per token, it achieves frontier-level performance in math, logic, and agentic coding while reaching up to 350 tokens/sec. Supports 256K context. NOTE: This model runs via StepFun, which may log and train on your prompts.
DiffusionGemma
google/diffusiongemmaDiffusionGemma is a high-speed diffusion-based version of Gemma 4 26B A4B. It supports optional reasoning and a 262,144-token context window.
Gemma 4 26B A4B Cybersecurity
google/gemma-4-26b-a4b-it-cybersecurityGemma 4 26B A4B Cybersecurity is a cybersecurity-focused variant based on the uncensored model, with provider moderation for illegal activities. It supports optional reasoning, image understanding, tool calling, and a 262,144-token context window.
Nemotron 3.5 Content Safety
nvidia/nemotron-3.5-content-safetyNemotron 3.5 Content Safety is a content moderation classifier that labels user messages and assistant responses as safe or unsafe. It supports a 131,072-token context window and optional reasoning.