ComfyUI Adds Wan2.7 Video Generation via Partner Nodes
Alibaba Wan2.7 video generation model is now available directly in ComfyUI through Partner Nodes, bringing comprehensive video creation capabilities to the popular node-based workflow tool.
Deep dives, tutorials, and analysis for AI-powered creators.
Alibaba Wan2.7 video generation model is now available directly in ComfyUI through Partner Nodes, bringing comprehensive video creation capabilities to the popular node-based workflow tool.
Tencent has released OmniWeaving, a unified video generation model that handles seven distinct tasks from text-to-video to reasoning-augmented generation, with publicly available weights and code.
Sony Interactive Entertainment has acquired Cinemersive Labs, a UK-based startup specializing in AI-powered technology that converts ordinary photos and videos into immersive 3D volumetric experiences.
OpenAI has launched ChatGPT Voice Mode on Apple CarPlay, giving drivers hands-free access to AI conversations for advice, brainstorming, and language practice while on the road.
Figma launched Make Kits and Make Attachments, two features that let design system authors teach Figma Make how to use their components and reference real data during AI generation.
OpenAI acquired TBPN, the Technology Business Programming Network, a live tech news show with 300,000 followers, for low hundreds of millions of dollars.
Cursor shipped a complete overhaul on April 2, replacing its code editor with a unified workspace built around AI agents, multi-agent management, and cloud-local handoff.
ElevenLabs launched ElevenMusic, a free iOS app that lets users create and remix songs using text prompts. The app marks the company expansion from voice AI into music generation.
Sony AI has released Woosh, an open source sound effects foundation model supporting text-to-audio and video-to-audio generation.
Google Vids now lets Workspace users create AI-directed avatars using Veo 3.1, generate custom soundtracks with Lyria 3, and produce up to 10 free video clips per month.
Microsoft launched three in-house AI models on April 2, reducing its dependence on OpenAI. MAI-Image-2, MAI-Voice-1, and MAI-Transcribe-1 cover image generation, text-to-speech, and speech recognition.
Alibaba releases Qwen3.6-Plus with 1M-token context, autonomous repository-level coding, and the ability to convert wireframes and screenshots directly into working frontend code.
GitHub launches /fleet in Copilot CLI, dispatching multiple AI agents to work on different tasks simultaneously instead of processing requests sequentially.
Liquid AI releases LFM2.5-350M, a 350M parameter open-weight model that runs AI agents on mobile GPUs in just 81MB with 32k context window.
Z.ai launches GLM-5V-Turbo, a vision-coding model that scored 94.8 on Design2Code versus 77.3 for Claude Opus 4.6. Converts mockups directly to frontend code.
Arcee AI releases Trinity-Large-Thinking under Apache 2.0, ranking #2 on PinchBench behind Claude Opus 4.6 at just $0.90 per million output tokens.
OmniVoice is a zero-shot text-to-speech model supporting 600+ languages with voice cloning, an Apache 2.0 license, and inference 40x faster than real-time.
Alibaba released Wan2.7-Image, a unified AI model that handles both image generation and editing in one system, with fine-grained personalization, color control, and a Pro variant capable of 4K output.
Technology Innovation Institute releases Falcon Perception, a compact 0.6B parameter vision model that beats Meta SAM 3 across six key benchmarks while running on consumer hardware.
PrismML emerged from stealth and open-sourced Bonsai 8B, a true 1-bit large language model that fits 8.2 billion parameters into just 1.15 GB of memory, roughly 14x more compressed than comparable 16-bit models.
Google launches Veo 3.1 Lite, a cost-effective video generation model available now via the Gemini API and Google AI Studio.
H Company released Holo3, a vision-language model optimized for GUI agents that sets new state-of-the-art on the OSWorld-Verified benchmark, beating proprietary models with only 3B active parameters.