ComfyUI v0.29.0 Adds HeyGen, GPT-5.6, and Gemma4 Nodes

ComfyUI v0.29.0 Adds HeyGen, GPT-5.6, and Gemma4 Nodes

ComfyUI tagged v0.29.0 on July 28, 2026, bundling new HeyGen avatar and voice nodes, OpenAI GPT-5.6, Google Gemini 3.5 Flash, Gemma4 12B, ByteDance seed-audio, and a JoyImageEdit editor in one release.

Rescript: Free Open-Source Descript Alternative

Rescript: Free Open-Source Descript Alternative

Rescript is a free, open-source alternative to Descript that lets you edit video and audio by editing the transcript text, running entirely in your browser with local Whisper transcription.

Hubo Adds Two-Agent Code Review to Claude Code

Hubo Adds Two-Agent Code Review to Claude Code

Hubo is an open-source plugin that adds a two-agent implement-and-review loop to AI coding tools like Claude Code and Codex, so every change is critiqued before it reaches you.

audio.cpp 0.4 Runs Higgs Audio v3 TTS Locally

audio.cpp 0.4 Runs Higgs Audio v3 TTS Locally

audio.cpp 0.4 adds Higgs Audio v3 4B, Fish Audio S2 Pro, and Voxtral Realtime ASR to one C++/ggml engine, running flagship text-to-speech locally on CUDA with no Python.

Echo Pools Open-Weight Models at 1/3 Claude's Cost

Echo Pools Open-Weight Models at 1/3 Claude's Cost

Echo, a new public-alpha endpoint from Tracer, pools open-weight models like GLM-5.2 and Kimi behind one OpenAI-compatible API, promising Claude-class output at roughly a third of the cost.

NvChat: Free Desktop Client for NVIDIA's LLM API

NvChat: Free Desktop Client for NVIDIA's LLM API

NvChat is a free, open-source Windows app that gives you a native desktop chat interface for NVIDIA free hosted LLM API, with access to more than 100 models including vision and reasoning models.

SynthCut: An Open-Source AI Video Editor Run via MCP

SynthCut: An Open-Source AI Video Editor Run via MCP

SynthCut is a free, GPL-3.0 video editor built to be operated by an AI agent, exposing itself as a Model Context Protocol server so a client like Claude Desktop can run real, frame-accurate edits locally.

Qwen-Image-Flash: NVIDIA's 4-Step Open Image Model

Qwen-Image-Flash: NVIDIA's 4-Step Open Image Model

NVIDIA released Qwen-Image-Flash, a four-step distilled version of Alibaba's 20B Qwen-Image that keeps about 96 percent of the quality while cutting denoising steps more than 12 times.

Microsoft Mage: 4B Open Image Model Beats Giants

Microsoft Mage: 4B Open Image Model Beats Giants

Microsoft's Mage Team released Mage-Flow, a 4B open-weights image generation and editing model that it says matches or beats open systems five to eight times its size, with a Turbo variant that renders 1024px images in 0.59 seconds.

Ditto: Clone Any Website Into Clean Next.js Code

Ditto: Clone Any Website Into Clean Next.js Code

Ditto is a free, open-source tool that clones any public website into clean, componentized front-end code in about five minutes. Unlike most AI cloners, it is fully deterministic.

ShotPlan: Multi-Shot Cinematic AI Video From One Prompt

ShotPlan: Multi-Shot Cinematic AI Video From One Prompt

ShotPlan is a new open framework from TeleAI and Harbin Institute of Technology that generates multi-shot cinematic video from one text prompt, placing frame-accurate cuts, cross-fades, and camera moves in a single pass.

Free Weekly Newsletter

Stay ahead of Creative AI

Join creators getting the latest AI tools, model releases, and workflow tips delivered weekly.

No spam. Unsubscribe anytime.