Smarter AI agents are easier to manipulate
Injecting malicious instructions into an AI planner can corrupt every downstream agent — and more capable models are more vulnerable.
Injecting malicious instructions into an AI planner can corrupt every downstream agent — and more capable models are more vulnerable.