OpenAI GPT-Live: Full-Duplex ChatGPT Voice

OpenAI GPT-Live: Full-Duplex ChatGPT Voice

OpenAI launched GPT-Live on July 8, 2026, a pair of full-duplex voice models that listen and speak at the same time, replacing Advanced Voice Mode for every ChatGPT user.

AMD Ryzen AI Halo: $3,999 Local AI Workstation

AMD Ryzen AI Halo: $3,999 Local AI Workstation

AMD put a 128GB local AI workstation on retail shelves for $3,999. The Ryzen AI Halo runs models up to 200 billion parameters and undercuts Nvidia's DGX Spark by $700.

Gladia CLI: Solaria Speech-to-Text From Your Terminal

Gladia CLI: Solaria Speech-to-Text From Your Terminal

Gladia shipped a command-line version of its speech-to-text platform, turning audio transcription into one terminal command with SRT and VTT subtitles, speaker diarization, and 100-plus languages.

Tencent Hy3: Open 295B Agentic Coding Model

Tencent Hy3: Open 295B Agentic Coding Model

Tencent open-sourced Hy3, a 295B mixture-of-experts model (21B active, 256K context) under Apache 2.0. It beats GLM-5.2 at roughly half the size and is free on OpenRouter until July 21.

LongCat-2.0: China's 1.6T Open-Weights Coding Model

LongCat-2.0: China's 1.6T Open-Weights Coding Model

Meituan open-sourced LongCat-2.0, a 1.6-trillion-parameter agentic coding model trained entirely on Chinese chips. The weights just landed on Hugging Face, and it beats GPT-5.5 on SWE-bench Pro.

First Open-Source Diffusion ASR Model: How It Works

First Open-Source Diffusion ASR Model: How It Works

Interfaze's diffusion-gemma-asr-small is billed as the first open-source diffusion-based speech recognition model, refining random tokens into a transcript instead of decoding left to right.

Free Weekly Newsletter

Stay ahead of Creative AI

Join creators getting the latest AI tools, model releases, and workflow tips delivered weekly.

No spam. Unsubscribe anytime.