RL training can silently break unseen tasks
Researchers show how RL fine-tuning can make AI models actively fail on tasks they were never trained on, even when those tasks are identical.
Researchers show how RL fine-tuning can make AI models actively fail on tasks they were never trained on, even when those tasks are identical.