AI Coding, One Year Later: What August 2025 Didn't See Coming

Throwback Thursday. A year ago the best coding model had 200K context and scored 49% on SWE-bench. Today Claude Fable 5 scores 95% with 1M context. Here's the gap model by model

Read Original

Related