StepFun · Step 5
StepFun's flagship 600B/27B-active MoE agentic reasoning model with 1M context and text/image/video input, API-only in preview with open weights promised for mid-October 2026.
Leader: Claude Opus 5.5 at 58
Leader: Claude Mythos 5.1 at 60.9%
Leader: Claude Opus 5.5 at 66.9%
Leader: Claude Opus 5.5 at 1846
Leader: Claude Opus 5.5 at 61.4%
Leader: GPT-5.6 Sol at 32.3%
Leader: Kimi K3 at 88.7%
Leader: Claude Opus 5.5 at 87.7%
| Result | Score | Reported by | Source |
|---|---|---|---|
| AA-Omniscience Index · AA | 16.4 index (-100..100) | independent | Artificial Analysis: Step 5 Preview ↗ |
| DeepSWE v1.1 | 67.7% | vendor | MarkTechPost: StepFun launches Step 5 Preview (StepFun-reported figures) ↗ |
| FrontierFinance · high effort | 66.4% | vendor | MarkTechPost: StepFun launches Step 5 Preview (StepFun-reported figures) ↗ |
| DRACO · high effort | 83.3% | vendor | MarkTechPost: StepFun launches Step 5 Preview (StepFun-reported figures) ↗ |
| StepCodeBench | 49% | vendor | MarkTechPost: StepFun launches Step 5 Preview (StepFun-reported figures) ↗ |
| ProgramBench · as reported by StepFun; metric may differ from other labs' ProgramBench figures | 80.5% | vendor | MarkTechPost: StepFun launches Step 5 Preview (StepFun-reported figures) ↗ |