AI agents often ignore security boundaries
A new benchmark reveals that AI agents frequently cross forbidden boundaries even when explicitly told not to.
A new benchmark reveals that AI agents frequently cross forbidden boundaries even when explicitly told not to.