Audio8 TTS: Open 0.6B Voice Cloning, 11 Languages
Audio8 TTS Preview 0.6B is an open-weights, Apache 2.0 text-to-speech model that clones a voice from a few seconds of audio and speaks 11 languages.
Deep dives, tutorials, and analysis for AI-powered creators.
Audio8 TTS Preview 0.6B is an open-weights, Apache 2.0 text-to-speech model that clones a voice from a few seconds of audio and speaks 11 languages.
Google added a system-wide voice input mode to the Gemini Mac app: long-press Fn to dictate, edit, and generate images by voice in any window.
Storeshot generates App Store and Google Play screenshots from a single config file, built so AI agents can drive it. Open source and MIT licensed.
Segue saves your working context in one AI assistant and reloads it in another using a short spoken-word handle, no copy-pasting required.
Perplexity has brought its Personal Computer AI agent to Windows, reaching more than a billion devices with local file and Microsoft 365 automation.
xAI has launched Grok Build Mode, an early beta that turns a single chat prompt into a working, publishable app, website, or game.
Rescript is a free, open-source alternative to Descript that lets you edit video and audio by editing the transcript text, running entirely in your browser with local Whisper transcription.
Hubo is an open-source plugin that adds a two-agent implement-and-review loop to AI coding tools like Claude Code and Codex, so every change is critiqued before it reaches you.
audio.cpp 0.4 adds Higgs Audio v3 4B, Fish Audio S2 Pro, and Voxtral Realtime ASR to one C++/ggml engine, running flagship text-to-speech locally on CUDA with no Python.
Echo, a new public-alpha endpoint from Tracer, pools open-weight models like GLM-5.2 and Kimi behind one OpenAI-compatible API, promising Claude-class output at roughly a third of the cost.