Real-SWE: Benchmarking AI models on private, real-world, enterprise codebases
Real-SWE, a new benchmark evaluating frontier AI models on private, real-world enterprise codebases licensed from real companies, was released with Fable 5.1 running in Claude Code posting the top res…