AI still fails at true flexible reasoning
A new benchmark reveals that AI models are far weaker at flexible, novel reasoning than standard tests suggest.
A new benchmark reveals that AI models are far weaker at flexible, novel reasoning than standard tests suggest.