FLUX 3: One Model for Image, Video, Audio, and Action
Black Forest Labs shipped FLUX 3 on July 23, 2026, a single multimodal model that generates images, generates video with synchronized audio, and predicts physical action, all in one architecture.
Black Forest Labs shipped FLUX 3 on July 23, 2026, a single multimodal model that generates images, generates video with synchronized audio, and predicts physical action, all in one architecture.
Runway shipped its Media Router on July 23, 2026, an intelligent routing layer that automatically picks the best image, video, or audio model based on whether you prioritize quality, speed, or cost.
Moto is a new AI video editor that generates images, video, and motion graphics as fully editable clips directly on the timeline, not in a separate app.
Neill Blomkamp's 13-minute Seedance 2.0 short Nightborne holds continuity where most AI films fall apart. A case study in the reference-locking pipeline he built, and what a solo creator can replicate.
ShotPlan is a new open framework from TeleAI and Harbin Institute of Technology that generates multi-shot cinematic video from one text prompt, placing frame-accurate cuts, cross-fades, and camera moves in a single pass.
Edit a video already playing? Decart Lucy 2.5. Refining a finished clip? Runway Aleph 2.0. A brand new shot? Seedance 2.5. No per-use cost? An open model in Velorn. Live-feed editing is now a third pillar of AI video.
Generate every clip, cut them on a real timeline, add captions, and export a finished AI short without leaving ComfyUI. A full step-by-step Velorn workflow, plus how to automate it with an AI agent.
LocalClip is a new local-first AI video clipper for Mac that turns long recordings into vertical short-form clips entirely on your own machine.
Turn a single product photo into a scroll-stopping video ad with AI. A four-step workflow using image-to-video, AI voiceover, and a free editor.
ByteDance is rolling out Seedance 2.5 this week: a 30-second continuous 4K clip, a beta long-video mode near three minutes, and up to 50 references, on Dreamina and CapCut.
claude-real-video is a local CLI that pulls scene-change frames, dedupes repeats, and transcribes audio so Claude, ChatGPT, or Gemini can actually watch a video, not just read its transcript.
Runway shipped Agent Skills on July 2, one-command actions inside its marketing agent that build an ad campaign, create a commercial, or localize your ads from a single prompt, live across all plans.
Apple Creator Studio's June 2026 update threads AI captions, edit detection, auto mask, and image generation through Final Cut Pro, Logic Pro, and Pixelmator Pro.
ByteDance Seedance 2.0 Mini and native 4K video are now selectable inside ComfyUI's existing Seedance nodes, with ready-made templates for text, frame, and reference workflows.
HappyHorse 1.1 lands in ComfyUI as a partner node that generates video with synchronized dialogue and sound effects in a single render pass.
ComfyUI shipped v0.26.0 with native nodes for Kling V3-Turbo, Luma Rays 3.2, Krea2, Qwen3-VL, Boogu-Image, SCAIL-2, HappyHorse 1.1, and a new advanced 3D loader, all in one release.
How animator Xindi Zhang built her Oscar-shortlisted thesis film in ComfyUI using style transfer, custom LoRAs trained on her own footage, and AI morphing.
Run ByteDance's Bernini-R video and image editing model entirely on your own GPU in ComfyUI. A step-by-step GGUF workflow for local, reference-guided edits with no cloud render fees.
LTX Director 2.0 turns Lightricks LTX 2.3 into a full timeline-based video editor inside ComfyUI, adding IC-LoRA, Retake Mode, and audio inpainting.
ComfyUI shipped v0.25.0 and v0.25.1 in three days, adding Kling V3-Turbo, Depth Anything 3, Tripo3D, and native 3D preview nodes.
Runway's Aleph 2.0 video model is now available inside Figma Weave, letting designers and editors direct clips up to 30 seconds frame by frame from Weave's node-based canvas.
xAI pushed Grok Imagine Video 1.5 to wide release on June 17, 2026, topping the image-to-video leaderboard with native audio at about $4.20 per minute, roughly 86 percent below Sora 2.
OpenAI is retiring Sora, turning the AI video race into a two-horse contest. Here is how Kling 3 and Runway compare on price, resolution, audio, and multi-shot control, plus a safe migration path off Sora.
AI video generation costs anywhere from half a cent to over forty cents per second in 2026. This guide compares Veo, Runway, Kling, Luma and Avataar Varya on real cost per second and shows where the savings are.