quick-post
Apr 28, 2026
NVIDIA released Nemotron 3 Nano Omni on April 28: a 30B open-weight model that handles text, image, video, and audio in one architecture, with joint audio-visual reasoning and multi-hour context.
ComfyUI
Apr 28, 2026
ComfyOrg announced ComfyStudio on April 28: a Los Angeles VFX and animation residency that pairs working artists with the ComfyUI team and releases every workflow open-source.
quick-post
Apr 28, 2026
Poolside released the first two models in its Laguna family on April 28: a 225B-parameter flagship called M.1 and an open-weight 33B model called XS.2 under Apache 2.0. Both are pitched at agentic, long-horizon coding work.
Deep Dive
Apr 28, 2026
SenseTime open-sourced SenseNova U1 on April 28, 2026, releasing three model variants under Apache 2.0. The architecture drops VAEs entirely, unifying image generation and text reasoning in one space.
Deep Dive
Apr 27, 2026
The complete creator guide to ComfyUI in 2026: setup, models, partner nodes, working workflows for image, video, audio, and 3D, the ecosystem, and managed-vs-self-host trade-offs.
Deep Dive
Apr 27, 2026
The complete creator reference to open-source AI in 2026: LLMs (DeepSeek V4, Kimi, Qwen), image (FLUX, Wan 2.7), video (Wan, Skywork, LTX), 3D (HY-World, TRELLIS), audio (Voicebox, OmniVoice, Darwin-TTS), all license-aware.
Deep Dive
Apr 27, 2026
ComfyUI v0.20.1 lands April 27 with SUPIR upscaling, RIFE and FILM frame interpolation, SAM 3.1 segmentation, plus 4K Veo and Kling and GPT-Image-2 partner nodes.
Deep Dive
Apr 24, 2026
DeepSeek V4 Preview ships MIT-licensed 1.6T and 284B MoE models with 1M context at sub-$0.30 per million output tokens. What it means for creators.
AI Video
Apr 23, 2026
Eyeline Labs released Vista4D, an open-source framework that reshoots existing video from new camera angles using a 4D point cloud plus video diffusion.
AI Tools
Apr 22, 2026
Inclusion AI released LLaDA2.0-Uni, a 16B MoE diffusion model that unifies text-to-image generation, instruction-based editing, and image understanding in one Apache-2.0 checkpoint.
AI Models
Apr 22, 2026
Alibaba released Qwen3.6-27B on April 22, 2026. The 27B dense open-weight model beats its larger 35B-A3B sibling on coding, agent, and vision benchmarks.
voice-ai
Apr 22, 2026
Xiaomi released MiMo-V2.5 on April 22, an 8B open-weights speech recognizer plus a three-model TTS series with prompt-built voice generation.
AI Tools
Apr 21, 2026
Alibaba's Taobao and Tmall algorithm team released Tstars-Tryon 1.0 on April 21, 2026, a commercial-scale virtual try-on system already serving millions of Taobao shoppers.
Audio
Apr 20, 2026
Jamie Pine's Voicebox brings seven TTS engines and voice cloning to your local machine. Free, open source, and entirely offline.
AI News
Apr 19, 2026
ComfyUI added Quiver Arrow 1.1 as a Partner Node on April 20, 2026, bringing structured SVG generation to open-source visual workflows for the first time.
Deep Dive
Apr 19, 2026
Moonshot AI's Kimi K2.6 scores 58.6 on SWE-Bench Pro, placing first above GPT-5.4 and Claude Opus 4.6. Full architecture breakdown, benchmark analysis, and what the open-weight license actually allows.
AI Models
Apr 19, 2026
Alibaba released Qwen3.6-Max-Preview on April 20, 2026, topping six programming benchmarks including SWE-benchPro. Here is what it means for creators building AI workflows.
3d-ai
Apr 19, 2026
A community port of Microsoft TRELLIS.2 brings image-to-3D mesh generation to Apple Silicon. No Nvidia GPU required, GLB output ready for Blender and Unity.
AI Tools
Apr 15, 2026
Researchers published Darwin-TTS on April 15, 2026, a text-to-speech model that adds emotional expression to AI voice without any training, fine-tuning, or new data.
Image Generation
Apr 14, 2026
NucleusAI released Nucleus-Image on April 14, a 17-billion parameter diffusion model that activates only 2 billion parameters per image -- cutting compute without sacrificing quality.
AI Video
Apr 14, 2026
HeyGen launched a developer platform with an open-source CLI, an HTML-to-video rendering framework called Hyperframes, and agent skills for Claude Code, Codex, and Cursor.
Deep Dive
Apr 13, 2026
ByteDance's Seedance 2.0 is now a native node in ComfyUI, bringing multimodal audio-video generation with 12 simultaneous reference inputs directly into the most popular open-source creative AI pipeline.
ComfyUI
Apr 13, 2026
ComfyUI v0.19.0 lands with Ace Step 1.5 XL music generation, Qwen 3.5 text generation, SeeDance 2.0 video nodes, and a faster Flux 2 decoder.
Open Source
Apr 12, 2026
llama.cpp release b8769 adds audio multimodal support for Qwen3-Omni and Qwen3-ASR models, bringing local speech recognition and audio understanding to consumer hardware.