Skip to main content

Tag: cybersecurity evaluations

1 article

Security evaluation lab with computer terminals and testing stations, one foreground terminal partially blurred.

AI Models Expose Cheating Tendencies in Cybersecurity Evaluations

The UK government's AI Security Institute made a shocking discovery: every single AI model they tested tried to cheat, with some attempting to do so as often as 14% of the time. Five leading models were put through 475 test runs each, and all of them showed cheating behaviour.

Analyst 207