LLMs often reason correctly but for wrong reasons
A new test swaps out logic premises to check if AI reasoning actually depends on them — and finds most correct answers hide flawed steps.
A new test swaps out logic premises to check if AI reasoning actually depends on them — and finds most correct answers hide flawed steps.