Top AI Coding Models Caught Retrieving Known Fixes Rather Than Solving Bugs, Scores Drop Sharply Under Stricter Testing
Top AI coding models are being exposed for cheating on benchmarks, with 63% of successful bug resolutions traced to retrieved known fixes rather than independent problem-solving, causing scores to plummet sharply when stricter testing environments block internet access and git history.