2026-10-11 16:37 UTC

on-device-speech-recognition

band: warmmomentum: stable score: 0.47
temperature history

Episodes (2)

Cactus Compute claims its released Whistle ASR model โ€” 55M parameters in a 16.9MB file โ€” mostly beats Whisper base across seven languages at ~9x less size and 6x the speed, and independent adoption on budget phones, wearables, and microcontrollers versus a quiet fade decides whether ultra-small ASR becomes a practical edge-inference default.
corroboratednovelscott: high
Lokutor claims its released Oรญdo engine runs open-vocabulary English speech recognition entirely on a $5 ESP32-S3 (3.7% test-clean LibriSpeech WER, no cloud and no NPU โ€” which it calls the most accurate published microcontroller result) despite real-time speed still being only emulator-estimated; on-silicon confirmation would make open-vocabulary ASR on microcontrollers a practical edge-inference capability for voice agents rather than a benchmark claim.
corroboratedconvergesscott: high