2026-10-11 16:37 UTC

text-to-speech

band: coolmomentum: stable score: 0.006
temperature history

Episodes (6)

Independent testing will determine whether ClearVoice can run OmniVoice multilingual text-to-speech and voice cloning fully offline on supported iPhones and iPads with acceptable quality, latency, and peak memory use.
expiredconvergesscott: medium
Independent use will determine whether Mac Out Loud provides a practical, fully local open-model text-to-speech workflow on macOS without cloud inference or analytics.
expiredknownscott: medium
NineNineSix claims its Apache-2.0 Gepard 1.0 model reaches 68.7 milliseconds median time-to-first-audio and 5.23% WER on one RTX 4090, which would make it a leading low-latency open TTS option on commodity GPU hardware.
expiredconvergesscott: medium
BreezeBlue claims its released Breeze-TTS-2 model delivers frontier-quality text-to-speech in a roughly 7GB locally runnable package, potentially expanding high-quality self-hosted voice generation.
expiredconvergesscott: medium
CrispASR’s maintainers claim their released single-binary C++ runtime can run multilingual speech-recognition and text-to-speech models locally across commodity systems, potentially simplifying self-hosted audio applications.
expiredknownscott: low
VoxGen’s maintainer claims the released Rust and Vulkan runtime makes local VoxCPM2 speech generation practical on AMD hardware without Python, PyTorch, or CUDA dependencies.
expiredknownscott: medium

Trajectory notes