2026-10-11 17:12 UTC

Independent replication and adoption will determine whether the proposed intelligence-per-watt metric produces reproducible, decision-useful comparisons of local AI models and inference hardware.

state: expiredheat: lowuncertainty: highconvergesscott: mediuminference-economics local-inference model-evaluation

What is this?

Researchers associated with Stanford’s Hazy Research propose “intelligence per watt” (IPW), defined as mean task accuracy divided by mean inference power draw, for comparing local language-model and hardware combinations. They report an end-to-end, cross-platform profiling harness supporting NVIDIA, AMD, and Apple Silicon, tested across more than 20 local models, eight accelerators, and roughly one million queries. The supplied results report substantial efficiency gains, but they primarily describe the authors’ own study and related coverage; they do not establish independent replication or broad external adoption.

Why it matters to Scott

IPW independently operationalizes Scott’s model-plus-harness and AI unit-economics positions by evaluating useful accuracy against measured power for specific model–hardware combinations. If independently replicated, it could become a practical selection signal for his hardware-aware local inference and gamepc workloads; for now, the supplied evidence establishes only the authors’ benchmark, not external validation or adoption.
ip:concept.ai-unit-economicsip:concept.model-plus-harness-benchmark-unitip:concept.capability-auditdev:concept.hardware-aware-local-inferencedev:project.gamepcradar:concept.local-inferenceradar:concept.model-evaluationradar:concept.inference-economicsradar:concept.inference-efficiencyradar:concept.ai-hardware
queries asked of Scott's wikis
  • local inference economics and hardware efficiency
  • decision-useful evaluation metrics for local models
  • cross-platform reproducible inference benchmarking
  • accuracy versus energy cost tradeoffs
  • local versus cloud workload routing economics
  • evaluation harnesses for model-hardware combinations

Measured heat

no measured readings yet — the hourly heat pass fills this in

How the heat travelled

no chain yet — the hourly chain pass fills this in

Evidence (2) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 reddit[2511.07885] Intelligence per Watt: Measuring Intelligence Efficiency of Local AI
LocalLLaMA
pscoutou3011
🟧 echo.paper ⭐The authors introduce intelligence per watt (IPW), defined as task accuracy per unit of power, and report evaluating 20+ local models, 8 accJon Saad-Falcon, Avanika Narayan, Hakki Orhun Akengin, J. Wes Griffin, Herumb Shandilya, Adrian Gamarra Lafuente, Medhya Goel, Rebecca Joseph, Shlok Natarajan, Etash Kumar Guha, Shang Zhu, Ben Athiwaratkun, John Hennessy, Azalia Mirhoseini, Christopher Ré——

Interpretation history

Decision trace