Mistral Small 4 Ships 119B Open Multimodal Model
Mistral AI released Mistral Small 4, a 119B mixture-of-experts model that unifies reasoning, multimodal vision, and coding under a single Apache 2.0 license.
Mistral AI released Mistral Small 4, a 119B mixture-of-experts model that unifies reasoning, multimodal vision, and coding under a single Apache 2.0 license.
NVIDIA launched the Nemotron Coalition at GTC, uniting Black Forest Labs, Cursor, Mistral AI, Perplexity, and four other labs to co-develop open frontier AI models.
Anthropic announced on March 13 that its full 1 million token context window is now generally available for Claude Opus 4.6 and Sonnet 4.6 at standard API pricing. The company eliminated the long-context surcharge that previously added up to 100% to the cost of requests exceeding 200,000 tokens
Meta has delayed the launch of its next-generation AI model, codenamed "Avocado," from March to at least May 2026. Internal testing revealed the model falls short of competitors from Google and OpenAI, landing somewhere between Google's Gemini 2.5 and Gemini 3 in performance
NVIDIA released Nemotron 3 Super on March 11, 2026, an open-source 120B-parameter hybrid Mamba-Transformer model that delivers 5x higher throughput than its predecessor for agentic AI workloads
Anthropic launched Claude Code Review on March 9, 2026, a multi-agent system that dispatches teams of AI reviewers on every pull request. The tool uses color-coded severity ratings and boosted thorough code reviews from 16% to 54% in internal testing
OpenAI launched GPT-5.4 on March 5, 2026, introducing native Computer Use mode that lets the model read screens and control mouse and keyboard inputs directly. The release also brings a 1M token context window, 33% fewer factual errors across benchmarks, and 47% token efficiency gains
OpenAI released GPT-5.3 Instant on March 3, 2026, replacing GPT-5.2 Instant as the default model in ChatGPT. The update reduces hallucinations by 26.8% on high-stakes queries, eliminates the preachy tone users complained about, and ships to all ChatGPT users including the free tier
Zhipu AI released GLM-5 on February 13, 2026, a 744 billion parameter open-source model trained entirely on Huawei Ascend chips without a single NVIDIA GPU
Apple just admitted it cannot build a competitive voice assistant alone. On February 12, 2026, the company announced a multi-year partnership with Google to rebuild Siri from the ground up using Gemini AI
Perplexity launched Model Council on February 5, 2026, a feature that runs the same query across multiple frontier AI models simultaneously and synthesizes a single, cross-validated answer. The feature is available to Perplexity Max subscribers at $200 per month