VoiceHop Brings Real-Time AI Voice Translation to Video

VoiceHop Brings Real-Time AI Voice Translation to Video

VoiceHop 2.0 brings sub-second, voice-preserving AI translation to any video, stream, or call as a browser extension. Here is what it does and how real-time translation compares to async dubbing tools.

audio.cpp 0.4 Runs Higgs Audio v3 TTS Locally

audio.cpp 0.4 Runs Higgs Audio v3 TTS Locally

audio.cpp 0.4 adds Higgs Audio v3 4B, Fish Audio S2 Pro, and Voxtral Realtime ASR to one C++/ggml engine, running flagship text-to-speech locally on CUDA with no Python.

OpenAI GPT-Live: Full-Duplex ChatGPT Voice

OpenAI GPT-Live: Full-Duplex ChatGPT Voice

OpenAI launched GPT-Live on July 8, 2026, a pair of full-duplex voice models that listen and speak at the same time, replacing Advanced Voice Mode for every ChatGPT user.

Gladia CLI: Solaria Speech-to-Text From Your Terminal

Gladia CLI: Solaria Speech-to-Text From Your Terminal

Gladia shipped a command-line version of its speech-to-text platform, turning audio transcription into one terminal command with SRT and VTT subtitles, speaker diarization, and 100-plus languages.

First Open-Source Diffusion ASR Model: How It Works

First Open-Source Diffusion ASR Model: How It Works

Interfaze's diffusion-gemma-asr-small is billed as the first open-source diffusion-based speech recognition model, refining random tokens into a transcript instead of decoding left to right.

The Best AI Music Generators in 2026: Suno, Udio, ElevenLabs and More

The Best AI Music Generators in 2026: Suno, Udio, ElevenLabs and More

The best AI music generators in 2026, from Suno and Udio to ElevenLabs Music, Google Lyria, and Stable Audio. We compare standout features, pricing, and commercial-use rights so you can pick the right tool for songs, scores, or royalty-free background music.

MOSS-Audio 8B: Open-Source Audio AI Beating 30B Rivals

MOSS-Audio 8B: Open-Source Audio AI Beating 30B Rivals

OpenMOSS published the MOSS-Audio technical report on June 1, 2026, documenting four open-source audio-language models that achieve benchmark scores rivaling systems three to four times their size.

Free Weekly Newsletter

Stay ahead of Creative AI

Join creators getting the latest AI tools, model releases, and workflow tips delivered weekly.

No spam. Unsubscribe anytime.