computer-use-agents
band: warmmomentum: stable
score: 0.263
Episodes (14)
Trajectory notes
- 2026-09-11T00:27:27Z: apple-safari-ai-agent-interface closed (faded) — Apple’s browser-native MCP tooling independently converges with Scott’s established “hands and eyes” architecture: coding agents need structured action plus DOM, console, network, screenshot and execution feedback to run reliable
- 2026-09-09T07:23:24Z: openai-cua-sample-app closed (faded) — OpenAI’s browser action/observation implementation converges with Scott’s AI Chrome Vision case study and OpenAI/Selenium loop in Scrape, supplying a concrete reference harness to compare against his browser-automation implementations. Tha
- 2026-09-07T14:34:45Z: codex-bundled-libreoffice-runtime closed (faded) — OpenAI’s reported bundling of LibreOffice is a consequential implementation of Scott’s “give the agent a workshop” and Code-First Architecture positions: system capability comes from packaging a capable execution substrate arou
- 2026-09-07T05:25:38Z: plugclaw-private-mobile-gui-agent closed (faded) — This is another mobile computer-use/privacy-appliance claim in territory already covered by Scott’s Agent Addressability and SiloOS frameworks and by the radar’s Android Remote Control MCP case. The GUI-control approach is the
- 2026-09-07T01:29:04Z: uia-reader-semantic-desktop-perception closed (faded) — The radar already tracks the same accessibility-tree-based desktop-agent approach in `radar:agent-desktop-accessibility-automation`, while Scott’s Denoised Semantic DOM and “Text Is The Model’s Home Turf” pages already arg
- 2026-09-03T20:29:48Z: cogram-studio-agentic-cad-bim closed (faded) — Cogram independently implements Scott’s agent-native/workshop position: a headless, agent-addressable engineering substrate operated through MCP while producing human-legible CAD/BIM artifacts and drawings. It also extends that arc
- 2026-08-31T23:35:48Z: simular-sai-osworld-cost-lead closed (faded) — Scott already treats agent performance as a model-plus-harness property and evaluates cost against validated capability, as captured in Model-Plus-Harness Benchmark Unit and Inference-Time Scaling; the radar also already tracks com
- 2026-08-29T18:30:09Z: cua-lite-local-computer-use-stack closed (faded) — The claimed stack independently packages several positions Scott already holds: computer-use capability is a model-plus-harness property, agents need observable hands-and-eyes environments, disposable sandboxes, and retained tr
- 2026-08-22T14:37:37Z: tencent-ui-mate-27b-validation closed (faded) — Scott already holds the governing position in Evaluation-Driven Development and Trace-backed agent comparison: claimed computer-use capability is not established until repeatable, end-to-end runs expose failures and completion evi