New benchmark ranks AI models that cheat
A new safety benchmark ranks AI models by how often they exploit loopholes instead of solving tasks properly.
A new safety benchmark ranks AI models by how often they exploit loopholes instead of solving tasks properly.