2026-10-11 16:38 UTC

Neuphonic claims its open-source NeuDecide β€” a 43MB model that maps audio directly to tool calls without transcription β€” enables practical voice-enabled agent workflows on edge devices via WASM browser deployment.

state: corroboratedheat: mediumuncertainty: lowconvergesscott: highaudio-to-tool edge-inference agent-harnesses voice-agentsNeuphonicTeamNeuphonic

What is this?

Neuphonic is a voice AI company releasing open-source, on-device speech models (NeuTTS Air, NeuCodec) built on small LLM backbones (Qwen 0.5B) for CPU-native inference. The web results confirm their TTS/voice-cloning models and LiveKit integration, but do not surface a model called 'NeuDecide' that maps audio directly to tool calls β€” the search returns only NeuTTS Air (text-to-speech) and NeuCodec. The case's central claim (a 43MB audio-to-tool model with WASM browser demo) is not corroborated by the supplied snippets; it may be a newer/ ΰ€…ΰ€²ΰ€— release not yet indexed, or the name may be conflated with NeuTTS Air.

Why it matters to Scott

A 43MB open-weight audio-to-tool model running in-browser via WASM is a concrete arrival at positions Scott's canon has argued for years: the fast lane of the Fast-Slow Split served by tiny local models (ip:framework.fast-slow-split), model sovereignty via open weights that run on edge hardware without vendor permission (ip:framework.sovereign-software-assurance, ip:concept.model-perishability), hardware-aware local inference as explicit runtime policy (dev:concept.hardware-aware-local-inference), and the shallow-action pool executed without round-tripping through ASR (ip:framework.voice-ais-fork). If the release holds up, it gives Scott a dated-receipts opportunity: a consequential other party shipping the architecture he specified.
ip:framework.fast-slow-splitip:framework.voice-ais-forkip:framework.sovereign-software-assuranceip:concept.model-perishabilitydev:concept.hardware-aware-local-inferenceip:framework.micro-agents-architecturedev:concept.padded-cell-agent-architecturedev:concept.deterministic-agent-control-planeip:framework.agent-native-computingdev:project.audioradar:aiope-android-agent-runtimeradar:1dial-real-world-task-agentradar:adaptive-kv-cache-streamingradar:aa-agentperf-local-benchmark
queries asked of Scott's wikis
  • audio-to-tool calling without ASR transcription
  • edge voice agent architectures WASM browser deployment
  • on-device function calling from raw audio
  • model sovereignty open-weight voice agents
  • local inference economics for voice agents
  • agent harness patterns for streaming audio input

Measured heat

now 0 pts/hpeak 4 pts/hcomments 0/hpeers p26momentum: steady2 platformsage 77h
points/hour across evidence Β· reading as of 2026-10-12 02:59:37.977291+11:00 Β· deterministic, not a model opinion

How the heat travelled

10-08 11:16⭐ origin directly observedWe’re open sourcing NeuDecide: a 43 MB audio-to-tool model with a WASM browser demo
TeamNeuphonic on r/LocalLLaMA
β€”
10-11 10:39first on hacker news Β· published Β· +71.4hVoice racer game built with OSS model NeuDecide
neuphonic_s
β€”
10-08 11:16amplified on r/LocalLLaMA πŸ‘‘reddit.post.1x0o8zr
TeamNeuphonic
peak 9 Β· 2 comments Β· 85% of case engagement
10-11 10:39amplified on hacker newshn.story.50041695
neuphonic_s
peak 1 Β· 0 comments Β· 15% of case engagement
10-08 12:31our radar first saw it Β· +1.2hdiscovery anchor: reddit.post.1x0o8zrβ€”
pace: p45 vs 1243 stories at the 72h mark (now 77h old) β€” ahead of apowerb-open-agent-runtime (1.2x), behind agentic-flooding-public-services (0.9x)

Evidence (2) β€” ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 reddit ⭐We’re open sourcing NeuDecide: a 43 MB audio-to-tool model with a WASM browser demo
LocalLLaMA
TeamNeuphonic92
🟧 hnVoice racer game built with OSS model NeuDecideneuphonic_s10

Interpretation history

Decision trace