Not All LLM Reasoners Are Created Equal
Arian Hosseini, Alessandro Sordoni, Daniel Toyama, et al.
This paper reveals a significant reasoning gap in LLMs when solving compositional math problems, showing that performance on standard benchmarks masks systematic differences in reasoning abilities.