Qwen3.5-Omni Handles Text, Image, Audio, Video in One Model
Alibaba releases Qwen3.5-Omni, a multimodal AI model that processes text, images, audio, and video while generating real-time speech output in 36 languages.
Alibaba releases Qwen3.5-Omni, a multimodal AI model that processes text, images, audio, and video while generating real-time speech output in 36 languages.
ByteDance is rolling out Dreamina Seedance 2.0 globally through CapCut after pausing the launch to address copyright concerns from Hollywood studios. The model now blocks real-face video generation.
Google launched Gemini 3.1 Flash Live, a new audio-first AI model designed for real-time voice conversations. It supports over 90 languages, filters background noise more effectively, and maintains context for twice as long.
Cohere released Transcribe, a 2-billion-parameter open-source speech recognition model that tops the HuggingFace Open ASR Leaderboard with a 5.42% average word error rate, beating OpenAI Whisper Large v3 by 27%.
Intel launches the Arc Pro B70 with 32GB GDDR6 VRAM and 367 INT8 TOPS at just $949, giving AI creators an affordable alternative to NVIDIA RTX Pro workstation GPUs.
ComfyUI introduces a custom PyTorch memory allocator called Dynamic VRAM that eliminates out-of-memory crashes and lets creators run large AI models on limited hardware.
xAI launched SuperGrok Lite, a $10-per-month subscription tier that includes basic AI image and video generation. The plan sits below the existing $30 SuperGrok tier and targets casual creators.
Udio settles lawsuits with Universal Music Group and Warner Music Group and pivots to a walled-garden fan platform for licensed remixing and mashups.
Mistral AI released Voxtral TTS, a 4-billion-parameter open-weights text-to-speech model that matches or beats ElevenLabs on naturalness benchmarks. It supports nine languages and clones voices from just three seconds of reference audio.
Meta is spending $600 billion on AI data centers, cutting 16,000 jobs, and still falling behind Google and OpenAI. In a single week in March 2026, the company delayed its flagship AI model, acquired two agent startups, and reportedly began planning the largest layoffs in its history
Every AI assistant you have used so far has one fundamental limitation: it forgets you exist the moment you close the tab. Perplexity wants to change that
The MacBook Pro M5 Max can run a 70-billion parameter language model at 18 to 25 tokens per second, generate a FLUX image 3.8 times faster than its predecessor, and fit everything in 128GB of unified memory without offloading to the CPU
Meta Platforms signed a three-year AI content licensing deal with News Corp worth up to $50 million per year. The agreement gives Meta access to News Corp's US and UK news archives to train AI models and power real-time responses in Meta AI products across Facebook, Instagram, WhatsApp, and Messe...