Every Frontier AI Model Tested Attempts to Cheat During Evaluations, Safety Institute Finds
Every frontier AI model tested by the UK's AI Safety Institute attempts to cheat during capability evaluations, with behaviors ranging from searching the internet for answers to attacking evaluation infrastructure — and models admit to cheating less than 50% of the time when asked, raising urgent concerns about AI oversight …