2026-10-11 18:01 UTC

Halv’s creator claims its desktop coding-agent workspace reduced tokens per correct answer by 51.1% across 20 paired Codex SWE-rebench tasks through context compression, output filtering, and repository indexing, potentially lowering coding-agent inference costs.

state: expiredheat: lowuncertainty: highknownscott: lowcoding-agents context-management inference-economicsHalvEnslavedFish

What is this?

Halv is presented in the case as a desktop coding-agent workspace whose creator, identified as EnslavedFish, claims a 51.1% reduction in tokens per correct answer across 20 paired Codex SWE-rebench tasks using context compression, output filtering, and repository indexing. Halv’s website snippet supports the repository-indexing and command-output-filtering features, but describes a different evaluation: 48 questions in Django and SymPy sessions, reporting a 47% reduction in cost per correct answer and 59% fewer tokens per correct answer for SymPy. The supplied snippets do not verify the creator’s identity or the exact 51.1% result; the other research results concern separate systems, not independent validation of Halv.

Why it matters to Scott

Halv repeats the selective-context position already held in Scott’s Context Engineering page and implemented through lossy compaction in Ask terminal agent; the radar also tracks comparable claims in tokencompress-agent-context-pruning, though no supplied hit tracks Halv itself. Its outcome-normalized savings could become useful evidence for evaluating Scott’s harness, but the creator-reported 51.1% result is unverified and differs from the website’s evaluation, so this currently adds another implementation example rather than a demonstrated reason to change what he builds or argues.
ip:framework.context-engineeringdev:project.askdev:concept.adaptive-source-context-compilationradar:tokencompress-agent-context-pruningradar:github-tool-output-cost-tradeoffradar:graphify-repository-map-context
queries asked of Scott's wikis
  • coding-agent harness context compression output filtering
  • repository indexing versus agent file exploration
  • agent evaluation paired tasks tokens per successful outcome
  • inference economics token savings cache pricing
  • context pruning information loss coding accuracy

Measured heat

no measured readings yet — the hourly heat pass fills this in

How the heat travelled

no chain yet — the hourly chain pass fills this in

Evidence (2) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 reddit ⭐I built a desktop workspace for Claude Code that cuts token usage by 51%
ClaudeAI
EnslavedFish6143
🟠 redditI built something to stop burning through my claude tokens so fast
ClaudeAI
DutyOnly430806

Interpretation history

Decision trace