audio.cpp 0.4 Runs Higgs Audio v3 TTS Locally

audio.cpp 0.4 Runs Higgs Audio v3 TTS Locally

audio.cpp 0.4 adds Higgs Audio v3 4B, Fish Audio S2 Pro, and Voxtral Realtime ASR to one C++/ggml engine, running flagship text-to-speech locally on CUDA with no Python.

OpenReader v3: Self-Hosted TTS for PDF, EPUB, DOCX

OpenReader v3: Self-Hosted TTS for PDF, EPUB, DOCX

OpenReader v3.0 converts PDF, EPUB, DOCX, TXT, and Markdown files into synchronized read-along sessions or exported audiobooks, with multiple TTS providers and Docker deployment.

Voxtral TTS: Open Weights Challenge ElevenLabs

Voxtral TTS: Open Weights Challenge ElevenLabs

Mistral Voxtral TTS is a 4B parameter open-weights model that matches ElevenLabs quality in human evaluations, with 3-second voice cloning across 9 languages and self-hosting support.

Free Weekly Newsletter

Stay ahead of Creative AI

Join creators getting the latest AI tools, model releases, and workflow tips delivered weekly.

No spam. Unsubscribe anytime.