2026-10-11 17:10 UTC

Blocks.ai claims CLI-based agent tools can avoid roughly 26,000 tokens of MCP schema overhead, making CLI access materially more context- and cost-efficient for large tool sets.

state: expiredheat: lowuncertainty: highconvergesscott: mediumagent-harnesses inference-economics mcp tool-useBlocks.ai

What is this?

Blocks.ai claims that exposing large tool sets to agents through MCP can consume roughly 26,000 tokens in tool-schema definitions before the agent processes the task, whereas CLI tools incur little or no upfront schema cost and reveal usage details on demand. The supplied snippets broadly support the underlying tradeoff: MCP pays upfront for structured tool definitions and may return tighter structured data, while CLI access preserves context at scale but can shift burdens such as output parsing, authentication, and session management onto the agent. The snippets do not independently establish Blocks.ai’s identity, methodology, or the exact 26,000-token measurement.

Why it matters to Scott

Blocks.ai independently quantifies the exact context-tax argument in Scott’s Code-First Architecture and “Why Code Execution Beats MCP,” creating a potential dated-receipts and benchmarking opportunity. The claimed 26,000-token figure could strengthen that argument, but its significance is limited until Blocks.ai’s methodology and measurement are independently established; the radar already tracks closely related MCP-schema compression work.
ip:framework.code-first-architectureip:source.why-code-execution-beats-mcpip:framework.context-engineeringradar:mcptoon-tool-discovery-compressionradar:tokencompress-agent-context-pruningradar:concept.mcp
queries asked of Scott's wikis
  • MCP schema overhead and context budgeting
  • CLI-first agent tool harnesses
  • on-demand tool discovery and progressive disclosure
  • tool-use token economics at scale
  • MCP versus shell tooling tradeoffs
  • agent handling of CLI output and authentication

Measured heat

no measured readings yet — the hourly heat pass fills this in

How the heat travelled

no chain yet — the hourly chain pass fills this in

Evidence (5) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟧 hnMCP vs. CLI: 26,000 tokens burnt before your agent reads the promptkayleykiwi21
🟧 echo.blog ⭐Claims MCP can burn 26,000 tokens before an agent reads its prompt and compares that overhead with CLI tool access.Blocks.ai——
🟠 redditI built only-cli: Turn any website into a compact CLI tailored for AI agents. Browse the web in hundreds of tokens, not tens of thousands.
ClaudeAI
GeekLifer19738
🟠 reddit20 bucks is 20 bucks. Claude uses tokens for your connectors.
ClaudeAI
iliadz10
🟠 redditI've been dealing with the MCP side for a while, and I wanted to share what finally came up: mcpify.
OpenAI
BeneficialPenalty58910

Interpretation history

Decision trace