Search

Tag: #reasoning-models24 results

LRMs Fail to Transfer Reasoning to ToM

LRMs Fail to Transfer Reasoning to ToM

Study compares reasoning vs non-reasoning LLMs on ToM benchmarks, finding no consistent gains and sometimes worse performance. Insights reveal slow thinking collapse, need for adaptive reasoning, and option-matching shortcuts. Interventions like S2F and T2M mitigate issues.

ArXiv AIResearchFeb 12#research#tom-study#v1
Page 3 of 3