StepFun Launches Step 5 Preview: A 600B-Total, 27B-Active MoE Model With 1M Context for Long-Horizon Agentic Work

2026年9月21日 06:37
站內 AI 整理稿

StepFun has released Step 5 Preview, its new flagship model for agentic work.The target workloads are software engineering, professional knowledge work, and finance.The main pitch is cost.StepFun team states the model delivers comparable intelligence at a substantially lower task cost.

That is the ‘Pareto frontier’ framing in the launch title.Is it deployable?Yes, as a hosted API and on the StepFun platform.Self-hosting waits for open weights.StepFun says open weights land on October 15, 2026.By simple arithmetic, 600B parameters need about 1.2 TB in BF16, before KV cache.

Plan for multi-GPU server hardware once weights ship.What StepFun Shipped Step 5 Preview is a sparse Mixture-of-Experts (MoE) model.It holds about 600B total parameters and activates about 27B per token.That is roughly 4.5% of the weights per token.

The official model documentation lists these specs: Model ID: step-5-preview Context window: 1M tokens Input: text, images, and video Output: text Reasoning effort: low, medium, and high Streaming, tool calling, JSON Mode, JSON Schema, and prompt caching On research tasks, StepFun team states the model coordinated 950 web fetches in a single agent action.

StepFun team also documents a Claude Code integration through its Step Plan.Architecture: Narrow and Deep StepFun did not widen the network.It stacked 92 Transformer layers in a narrow-deep layout, according to Pandaily.

The research team argues deeper stacks give longer paths for implicit multi-hop reasoning.This matters during long prefill, when agents search, run code, and read tool returns.Training leans on on-policy, long-horizon reinforcement learning.

StepFun cites bit-wise train and inference alignment across MoE routing.Other listed techniques include MTP-3 speculative decoding, FP8 MoE, and KV-cache offload.StepFun reports more than 3x end-to-end speedup for long-horizon RL.Interactive Explainer (function(){var f=document.

getElementById("mtp-step5-x");window.addEventListener("message",function(e){if(f&&e.source===f.contentWindow&&e.data&&e.data.mtpH){f.style.height=e.data.mtpH+"px";}});})(); Benchmarks: Company-Reported vs Independent Step 5 Preview ran at High effort, while rivals ran at Max.

StepFun reports these results, via RuntimeWire: BenchmarkStep 5 PreviewClaude Opus 5GPT-6 AstraFrontierFinance66.469.755DRACO83.387.676.8 On coding, StepFun reports 67.7 on DeepSWE v1.1, 49.0 on StepCodeBench, and 80.5 on ProgramBench.GPT-6 Astra and Claude Opus 5 stay ahead on all 3.

StepCodeBench is StepFun’s own benchmark.StepFun also ran 2 agent experiments lasting 24 hours each.In the first, the model tuned an H100 kernel to 508 TFLOPS, against 493 for Claude Opus 5.In the second, it raised Qwen3-30B-A3B on AIME24 from 53.3% to 60% through automated post-training.

The independent check comes from Artificial Analysis.It scores Step 5 Preview at 44 on its Intelligence Index.The median for reasoning models in a similar price tier is 24.It measured output at 99.8 tokens per second on StepFun’s API.

Pricing StepFun’s API list prices per 1M tokens: Token typePriceInput, cache miss$1.00Input, cache hit$0.05Output, including reasoning$2.70 Artificial Analysis puts the medians for comparable models at $1.88 input and $10.00 output.There is 1 catch.

The model generated 160M output tokens on the index run, against a 92M median.Verbose reasoning eats part of the per-token savings.Key Takeaways StepFun’s Step 5 Preview is a 600B-total, 27B-active MoE model.It offers a 1M-token context with text, image, and video input.API pricing is $1.

00 input and $2.70 output per 1M tokens.Artificial Analysis scores it 44 on its Intelligence Index.Open weights are scheduled for October 15, 2026.FAQ What is Step 5 Preview?It is StepFun’s flagship MoE model for agentic coding, knowledge work, and finance.Is Step 5 Preview open weight?Not yet.

StepFun schedules open weights for October 15, 2026.How large is the context window?1M tokens, per StepFun’s documentation.How much does it cost?$1.00 per 1M input tokens and $2.70 per 1M output tokens.Check out the Technical Details.All credit goes to the researcher of this project.

Also, feel free to follow us on Twitter and don’t forget to join our 150k+ML SubReddit and Subscribe to our Newsletter.Wait!are you on telegram?now you can join us on telegram as well.Need to partner with us for promoting your GitHub Repo OR Hugging Face Page OR Product Release OR Webinar etc.?

Connect with us The post StepFun Launches Step 5 Preview: A 600B-Total, 27B-Active MoE Model With 1M Context for Long-Horizon Agentic Work appeared first on MarkTechPost.

Related

相關文章

AIbaseAI Agent

阿寶上線「魚來」摸魚AI系統,將碎片化休息納入智能管理

據瞭解,「魚來」首批已釋放500尾魚苗。用戶只需對阿寶說出「魚來」,即可隨機召喚一條魚,從中選出自己的「命定之魚」,並設置每日摸魚提醒。整個流程被概括為三步:對阿寶說出指令(如「每天下午2點提醒我來摸魚」)、為任務命名並開啟通知權限、坐等每日準時提醒。

剛剛
鈦媒體AI Agent

剪映殺入AI視頻賽道,LibTV們將迎來衝擊波?

AI價值官2026.09.21 11:48 · 來自浙江全文4155字00:00 / 12:39單點工具的比拼已成過去,打通創作全鏈路、佈局下一代內容形態,正成為所有玩家共同的競爭方向。文 | AI價值官,作者丨星野,編輯丨美圻過去一年多,AI視頻平臺快速崛起,悄然改寫了視頻生產的底層邏輯。一站式創作模式持續向短劇、漫劇、商業廣告等場景滲透,傳統剪輯軟件的地位正在受到衝擊。對所有剪輯工具而言,向上遊延伸能力邊界、深度整合AI生成能力,已從功能迭代的可選項,變成必須回應的命題。

剛剛
AIbaseAI Agent

賈躍亭的 FF 一口氣發佈九款 EAI 機器人,最貴超 92 萬元

據官方介紹,FF“四核全智”開放生態由 EAI 大腦與開發者平臺、EAI 本體、行業生產力解決方案和 EAI 數據工廠組成。賈躍亭表示,FF 要構建的不是單一形態的全能機器人,而是由通用泛化的 EAI 大腦支撐多種本體形態、並通過 Agents、Skills 和真實場景數據持續進化的機器人生態。

剛剛
AIBaseAI Agent

阿寶上線「魚來」摸魚AI系統,將碎片化休息納入智能管理

據瞭解,「魚來」首批已釋放500尾魚苗。用戶只需對阿寶說出「魚來」,即可隨機召喚一條魚,從中選出自己的「命定之魚」,並設置每日摸魚提醒。整個流程被概括為三步:對阿寶說出指令(如「每天下午2點提醒我來摸魚」)、為任務命名並開啟通知權限、坐等每日準時提醒。

1 小時前
AIBaseAI Agent

賈躍亭的 FF 一口氣發佈九款 EAI 機器人,最貴超 92 萬元

據官方介紹,FF“四核全智”開放生態由 EAI 大腦與開發者平臺、EAI 本體、行業生產力解決方案和 EAI 數據工廠組成。賈躍亭表示,FF 要構建的不是單一形態的全能機器人,而是由通用泛化的 EAI 大腦支撐多種本體形態、並通過 Agents、Skills 和真實場景數據持續進化的機器人生態。

5 小時前6400
IT之家AI Agent

法拉第未來一口氣發佈九款配置 EAI 機器人,最貴超 92 萬元

作者:遠洋 責編:遠洋 評論: 9 月 20 日消息,法拉第未來(Faraday Future,簡稱 FF)於美國當地時間 9 月 19 日舉行了 919 FF EAI 機器人“四核全智”系列新品發佈會,發佈了 FF All-New Futurist、FF Master Mini、FX Aegis Hyper、FX Aegis Mega 及 FX Aegis Classic Ultra-W 五大型號、九款配置 EAI 機器人本體新品,以及 K-12 教育、科研、安防和巡檢四套行業生產力解決方案。

9 小時前