Deep Dive
Sep 7, 2026
Breeze TTS 2 is the best-sounding open-weights voice model of the four here and the one you are least likely to be allowed to use. In open voice AI, license and hardware decide the pick long before audio quality does.
AI
Aug 25, 2026
BreezeBlue has open-sourced Breeze TTS 2, a text-to-speech model that now ranks first among open-weights systems on the Artificial Analysis Speech Arena.
AI
Aug 10, 2026
Say It, a free and open-source macOS app, turns highlighted text into spoken audio and clones voices entirely on your own machine, with no cloud or accounts.
Deep Dive
Jul 31, 2026
audio.cpp 0.5 turns a pure C++ engine into a full local audio studio: expressive DramaBox TTS, Confucius4 cross-lingual voice transfer, music, and AMD ROCm support, all with no Python.
Audio
Jul 29, 2026
Audio8 TTS Preview 0.6B is an open-weights, Apache 2.0 text-to-speech model that clones a voice from a few seconds of audio and speaks 11 languages.
Audio
Jul 24, 2026
audio.cpp 0.4 adds Higgs Audio v3 4B, Fish Audio S2 Pro, and Voxtral Realtime ASR to one C++/ggml engine, running flagship text-to-speech locally on CUDA with no Python.
Deep Dive
Jul 14, 2026
audio.cpp 0.3 adds Supertonic 3, IndexTTS2, Irodori-TTS, and MOSS-TTS to one C++/ggml engine, running local text-to-speech and voice cloning with no Python and no cloud fees.
Audio
Jul 7, 2026
PocketTTS-RAVEN runs text-to-speech and voice cloning entirely in a browser tab, with no server, GPU, or account, at faster-than-realtime speed.
AI
Jun 14, 2026
TTS Audio Suite shipped v5.0.0 for ComfyUI, adding Higgs Audio v3 voice cloning, Transformers 5, and Runtime Isolation for legacy engines.
Audio
Jun 8, 2026
A revamped open-source TTS benchmark now compares 46 text-to-speech models using objective scores and blind human voting, so creators can see which voices actually hold up.
tts
May 15, 2026
Supertonic 3 is an open-weights, CPU-only TTS engine from Supertone with 31 languages, expression tags, and zero-shot voice cloning.
Open Source
May 13, 2026
OpenReader v3.0 converts PDF, EPUB, DOCX, TXT, and Markdown files into synchronized read-along sessions or exported audiobooks, with multiple TTS providers and Docker deployment.
ai-voice
Apr 18, 2026
xAI launched the Grok Voice Agent API April 18 with standalone speech-to-text and text-to-speech endpoints priced at $0.10 per hour batch STT and $4.20 per million characters TTS.
Deep Dive
Apr 5, 2026
The AI voice cloning market reaches $4.06 billion in 2026. We compare ElevenLabs, Voxtral TTS, and Fish Audio S2 on quality, pricing, latency, and self-hosting to help creators choose the right tool.
Deep Dive
Mar 26, 2026
Mistral Voxtral TTS is a 4B parameter open-weights model that matches ElevenLabs quality in human evaluations, with 3-second voice cloning across 9 languages and self-hosting support.