Why AI safety rules can't be learned from data
A theoretical paper argues that safety constraints like 'don't escape the sandbox' are structurally invisible to AI training data.
A theoretical paper argues that safety constraints like 'don't escape the sandbox' are structurally invisible to AI training data.