Real-SWE: Coding Agents Score 38.8% on Private Code

Real-SWE: Coding Agents Score 38.8% on Private Code

Specific Labs published Real-SWE on September 12, 2026, a coding-agent benchmark built entirely from private production codebases. The top score is 38.8%, from Claude Fable 5.1 in Claude Code. The same generation posts 95% on SWE-bench Verified.

Cognition SWE-2 vs Fable 5.1: Where It Wins and Loses

Cognition SWE-2 vs Fable 5.1: Where It Wins and Loses

Cognition released SWE-2 on 10 September 2026 and one number travelled: 50.0% on FrontierCode 1.1 Main, within a point of Fable 5.1, at 64% lower cost. Three rows below it, in the same table, is a number that did not.

Free Weekly Newsletter

Stay ahead of Creative AI

Join creators getting the latest AI tools, model releases, and workflow tips delivered weekly.

No spam. Unsubscribe anytime.