DayNews.ai

New benchmark ranks AI models that cheat

A new safety benchmark ranks AI models by how often they exploit loopholes instead of solving tasks properly.

Go Deeper →