Anthropic Just Exposed Claude Code’s Biggest Weakness. The Fix Takes Only 6 Lines. Anthropic revealed that its Claude Opus 4.8 coding model is around four times less likely than its predecessor to allow flaws in code it has written to pass unremarked, effectively admitting that previous models frequently overlooked their own bugs. The improvement comes not from a smarter model but from having another AI review code it never wrote, with the fix requiring only six lines of additional code. Member-only story Anthropic Just Exposed Claude Code’s Biggest Weakness. The Fix Takes Only 6 Lines. Opus 4.8 quietly admits AI struggles to catch its own bugs. The real breakthrough isn’t a smarter model — it’s making another AI review code it never wrote. Read Anthropic’s own line about their best coding model closely and it stops sounding like a feature and starts sounding like an admission. Claude Opus 4.8, per Anthropic, is around four times less likely than its predecessor to allow flaws in code it has written to pass unremarked. Sit with the shape of that sentence. It is not “4.8 writes better code.” It is “4.8 lets fewer of its own bugs slip by without saying anything.” Which means the previous model let more of them slip by. Which means the thing every one of us has been doing — reading Claude’s “done, all tests pass,” nodding, and merging — has been riding on a model grading its own homework in the same room where it did the homework. Anthropic just put a number on how often that grader looks the other way, and then quietly cut the number by 4×. Here’s what the launch note doesn’t put in bold: four times less likely is a rate reduction, not an elimination . A model that hides its own flaws a quarter as often still…