Deep Dive
Sep 11, 2026
ElevenLabs released Music v2.5 on 11 September 2026 and made one decision that matters more than the model: lossless downloads on every plan, including Free, which gets five a day. Two days after Suno shipped v6, the AI music fight has moved from output quality to export rights.
Deep Dive
Sep 9, 2026
Ten AI tools now cover the whole podcast production line, and the cheap end got good enough that most shows should pay for exactly one seat. Every price verified against the vendor's own pricing page on September 9, 2026.
Deep Dive
Sep 8, 2026
On September 8, 2026, European lab Desert Ant Labs published 18 small AI models that run entirely on a users device, with SDKs for Swift, Kotlin and JavaScript and a free tier covering the first 100,000 monthly active devices per platform.
AI
Sep 2, 2026
TONE3000 launched a free, open-source plugin that streams more than 700,000 Neural Amp Modeler captures and impulse responses straight into your DAW.
AI
Sep 1, 2026
Meta shipped Muse Voice Transcribe on September 1, 2026, a real-time speech-to-text model that streams transcription, diarizes 20-plus speakers, and posts a 3.1% word error rate.
ai-music
Aug 31, 2026
Mureka released V9.5 of its AI music model, shifting focus from convincing sounds toward complete songs with stronger arrangements and more natural vocals.
AI
Aug 26, 2026
Google launched Gemini 3.5 Transcribe, a speech-to-text model that detects 85+ languages, removes filler words, and reaches a 2.6% word error rate.
AI
Aug 25, 2026
BreezeBlue has open-sourced Breeze TTS 2, a text-to-speech model that now ranks first among open-weights systems on the Artificial Analysis Speech Arena.
AI
Aug 20, 2026
ElevenLabs archived its local MCP server repository on August 20, 2026, pointing developers to a hosted server that authenticates with OAuth rather than a copied API key.
AI
Aug 10, 2026
Say It, a free and open-source macOS app, turns highlighted text into spoken audio and clones voices entirely on your own machine, with no cloud or accounts.
Deep Dive
Aug 10, 2026
NVIDIA released Magpie TTS Multilingual, an open-weights text-to-speech model that speaks 12 languages and hits a 32ms time-to-first-audio, small enough to self-host inside a live voice agent.
Deep Dive
Aug 10, 2026
Vocal Slice cuts audio by selecting words in a transcript, running Whisper locally so your recordings never leave your device.
Deep Dive
Aug 9, 2026
Vibez is a new open-source DAW written in pure Rust, launched August 9, 2026 for Linux, macOS, and Windows, pairing clip launching with a timeline plus VST3 and CLAP plugin support.
Deep Dive
Jul 31, 2026
audio.cpp 0.5 turns a pure C++ engine into a full local audio studio: expressive DramaBox TTS, Confucius4 cross-lingual voice transfer, music, and AMD ROCm support, all with no Python.
Deep Dive
Jul 29, 2026
GPT Transcribe ranked against Whisper, gpt-4o-transcribe, ElevenLabs Scribe, Gemini 3 Pro, and Voxtral on accuracy, price, and the features it drops.
Audio
Jul 29, 2026
Audio8 TTS Preview 0.6B is an open-weights, Apache 2.0 text-to-speech model that clones a voice from a few seconds of audio and speaks 11 languages.
Deep Dive
Jul 28, 2026
VoiceHop 2.0 brings sub-second, voice-preserving AI translation to any video, stream, or call as a browser extension. Here is what it does and how real-time translation compares to async dubbing tools.
Deep Dive
Jul 28, 2026
Fish Audio raised $52M and made S2.1 Pro free through August 31: production-grade voice cloning across 83 languages at roughly 90ms latency, plus an open-source path via Fish Speech.
Audio
Jul 24, 2026
audio.cpp 0.4 adds Higgs Audio v3 4B, Fish Audio S2 Pro, and Voxtral Realtime ASR to one C++/ggml engine, running flagship text-to-speech locally on CUDA with no Python.
Deep Dive
Jul 22, 2026
Alibaba's Tongyi Lab released Qwen-Audio-3.0-TTS, a hosted text-to-speech model that topped the Artificial Analysis TTS leaderboard in July 2026.
Deep Dive
Jul 14, 2026
audio.cpp 0.3 adds Supertonic 3, IndexTTS2, Irodori-TTS, and MOSS-TTS to one C++/ggml engine, running local text-to-speech and voice cloning with no Python and no cloud fees.
Deep Dive
Jul 13, 2026
Kyutai and Mirelo released MuScriptor, an open-weight model that transcribes a full song mix into separate, editable MIDI tracks, one per instrument.
Deep Dive
Jul 9, 2026
aria is a dependency-free native runtime that runs the full Stable Audio 3 text-to-music pipeline on ordinary GPUs, CPU-only laptops, and an 8GB Raspberry Pi 5, no Python required.
AI
Jul 8, 2026
Willow launched two speech-to-text models: Frontier Mini, a free unlimited tier, and Frontier Pro, a faster paid model built for power users.