2026-10-11 17:12 UTC

Independent evaluations will determine whether DeepSeek V4 Pro 0813 offers capability, latency, or price-performance advantages sufficient to change frontier-model selection for production workloads.

state: expiredheat: lowuncertainty: highknownscott: mediumdeepseek model-releases inference-economicsDeepSeekOpenRouter

What is this?

DeepSeek V4 Pro 0813 is presented as an open-weight, MIT-licensed mixture-of-experts model from DeepSeek, aimed at reasoning, software engineering, and long-running agentic workloads, with a lighter V4 Flash counterpart. The supplied evaluations agree that its substantially lower inference pricing could make it attractive for production model routing, but conflict on capability: CAISI places it roughly eight months behind the frontier, while other reviewers describe it as near or equal to leading closed models on selected benchmarks. Reported latency and price-performance advantages are workload- and configuration-dependent, so independent production evaluations remain necessary; the snippets provide only indirect attribution to OpenRouter.

Why it matters to Scott

Scott already holds the relevant position in Model Perishability and Capability Audit: models should remain swappable and be re-evaluated on representative production workloads rather than selected from launch benchmarks. V4 Pro is nevertheless an actionable candidate for his LiteLLM-based task-aware routing and trace-backed comparisons, while the radar already tracks the companion release on “deepseek-v4-flash-validation” and “deepseek-v4-flash-harness-efficiency.”
ip:concept.model-perishabilityip:concept.capability-auditip:concept.model-plus-harness-benchmark-unitdev:concept.task-aware-model-routingdev:concept.trace-backed-agent-comparisondev:technology.litellmradar:deepseek-v4-flash-validationradar:deepseek-v4-flash-harness-efficiencyradar:concept.inference-economicsradar:concept.model-routing
queries asked of Scott's wikis
  • model routing by capability latency and cost
  • production evals for coding and agentic models
  • open-weight inference economics and self-hosting
  • frontier quality versus price-performance
  • benchmark validity for long-horizon agent reliability
  • multi-model harnesses and workload-specific selection

Measured heat

no measured readings yet — the hourly heat pass fills this in

How the heat travelled

no chain yet — the hourly chain pass fills this in

Evidence (15) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟧 hnDeepSeek V4 Pro 0813explosion-s1038443
🟧 hnDeepSeek V4 Pro 0813 quietly releasedHiPHInch786
🟧 echo.other ⭐The official DeepSeek Models & Pricing page lists the current model version as “DeepSeek-V4-Pro-0813” (and Flash as “DeepSeek-V4-Flash-0731”DeepSeek——
🟧 hnDeepSeek v4 pro 0813 releasedalexwwang10
🟠 redditDeepSeek 0813 "Pro" vs GLM 5.2 & Kimi K3 🐋
singularity
VexObserver162
🟠 redditDeepSeek V4 Pro 0813 leaps over GLM-5.2
singularity
RetiredApostle14831
🟠 redditDeepSeek-V4-Pro-0813 released on api
LocalLLaMA
AlbeHxT978
🟠 redditDeepSeek: We’re launching DeepSeek-V4-Pro today!
LocalLLaMA
Nunki08509113
🟠 redditdeepseek-ai/DeepSeek-V4-Pro-0813 · Hugging Face
LocalLLaMA
mossy_troll_8452486
🟠 redditTested DeepSeek V4 Pro (0813) on coding with OpenCode & agentic work
LocalLLaMA
curiousily_017
🟠 redditdeepseek-ai/DeepSeek-V4-Pro-0813 (Available again) · Hugging Face
LocalLLaMA
panchovix924
🟠 redditSame demo, two failures on DeepSeek V4 Pro 0813, then V4 Flash finished it
artificial
neverontime511
🟠 redditThe DeepSeek V4 Pro 0813 frontend feels like Flash with more depth. The price decides whether it's worth it
singularity
Adventurous_Rush147401
🟧 hnDeepSeek-V4-Pro outperforms Fable 5 after fixing runtime inference controlDarenWatson71
🟧 hnDeepSeek-V4-Pro-0813 outperforming Fableed_mercer10

Interpretation history

Decision trace