Making AI agents safe is still unsolved
A review of 38 studies finds no current method can make AI agents reliably safe when they take real-world actions.
A review of 38 studies finds no current method can make AI agents reliably safe when they take real-world actions.