Tencent Hy3: Open 295B Agentic Coding Model

Tencent Hy3: Open 295B Agentic Coding Model

Tencent open-sourced Hy3, a 295B mixture-of-experts model (21B active, 256K context) under Apache 2.0. It beats GLM-5.2 at roughly half the size and is free on OpenRouter until July 21.

LongCat-2.0: China's 1.6T Open-Weights Coding Model

LongCat-2.0: China's 1.6T Open-Weights Coding Model

Meituan open-sourced LongCat-2.0, a 1.6-trillion-parameter agentic coding model trained entirely on Chinese chips. The weights just landed on Hugging Face, and it beats GPT-5.5 on SWE-bench Pro.

Claude Fable 5 Is Now Available: A Builder's Guide

Claude Fable 5 Is Now Available: A Builder's Guide

Claude Fable 5 is generally available again as of July 1, 2026: Anthropic's most capable widely released model for reasoning and long-horizon agents. What it is, the pricing, and the refusal-and-fallback change every builder must handle.

Claude Sonnet 5: Near-Flagship Power at a Lower Price

Claude Sonnet 5: Near-Flagship Power at a Lower Price

Anthropic launched Claude Sonnet 5 on June 30, 2026 as the new default model for Free and Pro, positioned close to Opus 4.8 at a lower price. Here is what changed, the tokenizer catch that resets your cost math, and how to switch your workflow.

DAI Studio Brings Visual Context Engineering to LLMs

DAI Studio Brings Visual Context Engineering to LLMs

Context engineering just got a dedicated workbench. On June 23, 2026, DAI Studio launched a free visual tool that lets you design, see, and version the exact context you feed into a large language model before it runs.

Sakana Fugu: One API to Orchestrate Top AI Models

Sakana Fugu: One API to Orchestrate Top AI Models

Sakana Fugu is a multi-agent system that behaves like one model: send a request to a single OpenAI-compatible API and it routes across frontier models for you. Here is what it costs, how it scores, and why orchestration matters for builders.

Nemotron 3 Ultra: NVIDIA's 550B Open-Weights MoE

Nemotron 3 Ultra: NVIDIA's 550B Open-Weights MoE

NVIDIA released Nemotron 3 Ultra on June 1 2026: a 550B mixture-of-experts model with 55B active parameters, open weights on Hugging Face, with 5x faster inference and 30% lower cost than Nemotron 2.

Qwen3.7-Max vs Claude, Gemini, GPT-5.5: Compared

Qwen3.7-Max vs Claude, Gemini, GPT-5.5: Compared

Qwen3.7-Max sets a new bar for non-hallucination rate among frontier agent models. Here is how it stacks up against Claude Opus 4.7, Gemini 3.1 Pro, and GPT-5.5 across the four reliability dimensions that decide which model goes into production.

NVIDIA Nemotron Diffusion: 3x Faster LLM Decoding

NVIDIA Nemotron Diffusion: 3x Faster LLM Decoding

NVIDIA released the Nemotron-Labs-Diffusion family on Hugging Face, an open-weights LLM that switches between autoregressive, diffusion, and self-speculation decoding for 2.7x to 3.3x throughput gains.

Free Weekly Newsletter

Stay ahead of Creative AI

Join creators getting the latest AI tools, model releases, and workflow tips delivered weekly.

No spam. Unsubscribe anytime.