Scientific research in LLM and LRM

Content:

Original link: The Illusion of Thinking: Understanding the Strengths and Limitations of Reasoning Models via the Lens of Problem Complexity / Apple Machine Learning Research.

The conclusion:

By comparing LRMs with their standard LLM counterparts under equivalent inference compute,
we identify three performance regimes: (1) low-complexity tasks where standard models surprisingly
outperform LRMs, (2) medium-complexity tasks where additional thinking in LRMs demonstrates
advantage, and (3) high-complexity tasks where both models experience complete collapse. 

Comments: