2026-10-11 16:38 UTC

Redditor No-Head-Royal, citing Artificial Analysis, reports that StepFun's released Step 5 Preview matches Kimi K3 (max)'s intelligence score of 44 at roughly one-third the price, potentially lowering the cost of accessing that measured capability tier.

state: corroboratedheat: mediumuncertainty: mediumconvergesscott: mediumfrontier-models inference-economicsStepFunArtificial AnalysisMoonshot AI
Surfaced 2026-09-21T23:25:26Z — Artificial Analysis lists Step 5 Preview as a StepFun model released September 18, 2026, scoring 44 on its Intelligence Index at $0.71 per t — Cross-platform attention now warrants high heat, but announcement quotations and discussion of purported weights still do not constitute independent validation of the price–quality claim. Reported architecture details sharpen the deployment questions, while the demo critique cautions against inferring coding-agent performance without establishing a benchmark contradiction.

What is this?

StepFun, a previously non-frontier Chinese AI lab, released Step 5 Preview around September 18–20, 2026. Artificial Analysis lists it on its Intelligence Index at 44 — the same score AA's own Kimi K3 (max) page shows — at roughly one-third Kimi's token price ($1.00/$2.70 per 1M input/output vs. K3's $3.00/$15.00), placing it on AA's cost–capability Pareto frontier. HN discussion quoting the official announcement describes a sparse MoE (600B total / 27B active) with 1M context and vision input, and StepFun has shipped Step Code, an MIT-licensed first-party coding-agent CLI built around the model. Caveats the snippets themselves raise: several other sources report Kimi K3 at 57 on the AA index (likely a different index version or configuration), Step 5 Preview's weight-access and license are listed as 'Pending' by third-party trackers after an apparently accidental early fork of BF16 weights (now 404, official weights promised October 15), and a demo video was credibly critiqued as reusing an existing project — so the capability-parity claim rests on one evaluator's snapshot, not independent replication.

Why it matters to Scott

StepFun independently demonstrates Scott's model-perishability/capability-symmetry thesis with dated receipts — a new Chinese lab reaching Kimi K3's measured tier at one-third price — and its MIT Step Code CLI continues the Chinese-lab first-party-open-harness lineage (Z.ai, MiniMax, DeepSeek) already on the radar. It bears on what he'd build: a direct candidate for his LiteLLM cost tiers if parity survives independent evaluation, a local-inference candidate when the promised Oct 15 weights land (27B-active MoE vs gamepc's constraints), and the unexplained 44-vs-57 AA index discrepancy is live material for his version-bound-assessment position.
ip:concept.model-perishabilityip:concept.capability-symmetrydev:concept.cost-tiered-llm-routingdev:technology.litellmdev:concept.hardware-aware-local-inferencedev:concept.version-bound-ai-assessmentradar:concept.inference-economicsradar:concept.open-weight-modelsradar:concept.agent-harnessesradar:concept.mixture-of-expertsradar:kimi-k3-single-consumer-gpuradar:zcode-open-runtime-release
queries asked of Scott's wikis
  • model routing tiers cost per task intelligence threshold
  • open-weights Chinese labs frontier parity sovereignty
  • first-party coding agent CLI pattern model-owned harness
  • sparse MoE local inference hardware requirements deployment
  • benchmark index versioning evaluator snapshot reliability
  • Pareto frontier price collapses model perishability

Measured heat

now 0 pts/hpeak 17 pts/hcomments 0/hpeers p25momentum: steady3 platformsage 578h
points/hour across evidence · reading as of 2026-10-12 02:59:37.977291+11:00 · deterministic, not a model opinion

How the heat travelled

09-17 14:00⭐ origin echo-reconstructedArtificial Analysis lists Step 5 Preview as a StepFun model released September 18, 2026, scoring 44 on its Intelligence Index at $0.71 per t
Artificial Analysis on other (echo) · attributed from reddit.post.1wkd8hx
—
09-19 05:18first on r/singularity · published · +39.3hStepFun joins the frontier: a previously non-frontier Chinese lab (StepFun) released a Kimi K3-level model, 3 times cheaper per Artificial Analysis
No-Head-Royal
—
09-19 05:42first on hacker news · published · +39.7hStepfun Step 5 Preview (LLM): On AA Pareto frontier
AnodicElegy
—
09-20 03:57first on r/LocalLLaMA · published · +62.0hthis looks promising: stepfun-ai/Step-5-Preview-BF16 · Hugging Face
adefa
—
09-19 05:18amplified on r/singularityreddit.post.1wkd8hx
No-Head-Royal
peak 87 · 12 comments · 11% of case engagement
09-19 05:42amplified on hacker newshn.story.49763660
AnodicElegy
peak 15 · 2 comments · 3% of case engagement
09-20 03:57amplified on r/LocalLLaMAreddit.post.1wl6gro
adefa
peak 52 · 33 comments · 9% of case engagement
09-20 04:35amplified on hacker news 👑hn.story.49772532
nateb2022
peak 140 · 33 comments · 34% of case engagement
09-20 11:47amplified on r/LocalLLaMAreddit.post.1wleycb
External_Mood4719
peak 33 · 16 comments · 5% of case engagement
09-23 15:12amplified on hacker newshn.story.49817387
AnneWodell
peak 1 · 1 comments · 0% of case engagement
2 more amplifiers in ainews.case_chain
09-19 05:20our radar first saw it · +39.3hdiscovery anchor: reddit.post.1wkd8hx—
09-21 23:25reached heat=high · +105.4h · via ledger——
pace: p82 vs 1032 stories at the 336h mark (now 578h old) — ahead of deathray-macos-webgpu-hang (1.0x), behind specific-real-swe-release (1.0x)

Evidence (9) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 redditStepFun joins the frontier: a previously non-frontier Chinese lab (StepFun) released a Kimi K3-level model, 3 times cheaper per Artificial Analysis
singularity
No-Head-Royal8712
🟧 echo.other ⭐Artificial Analysis lists Step 5 Preview as a StepFun model released September 18, 2026, scoring 44 on its Intelligence Index at $0.71 per tArtificial Analysis——
🟧 hnStepfun Step 5 Preview (LLM): On AA Pareto frontierAnodicElegy152
🟠 redditthis looks promising: stepfun-ai/Step-5-Preview-BF16 · Hugging Face
LocalLLaMA
adefa5133
🟧 hnStep 5 Preview: Advancing the Pareto Frontiernateb202214033
🟠 redditrene98c/Step-5-Preview-BF16 • HuggingFace (Fork)
LocalLLaMA
External_Mood47193315
🟧 hnStep Code: MIT-licensed coding agent CLI built around Step 5 PreviewAnneWodell11
🟧 hnStep 5 Preview, a 1M-context MoE from StepFun, shows up on OpenRouterAnneWodell12631
🟠 redditStepFun Step 5 preview showing up in openrouter
LocalLLaMA
cafedude3913

Interpretation history

Decision trace