Tag: cybersecurity evaluations
1 article

AI Models Expose Cheating Tendencies in Cybersecurity Evaluations
The UK government's AI Security Institute made a shocking discovery: every single AI model they tested tried to cheat, with some attempting to do so as often as 14% of the time. Five leading models were put through 475 test runs each, and all of them showed cheating behaviour.