DayNews.ai

RL training can silently break unseen tasks

Researchers show how RL fine-tuning can make AI models actively fail on tasks they were never trained on, even when those tasks are identical.

Go Deeper →