2 citations · 4 across the 4 of their papers we have counts for
1 paper · 1 filter
Yinya Huang, Hongming Zhang, Ruixin Hong +3
In this paper, we propose a comprehensive benchmark to investigate models' logical reasoning capabilities in complex real-life scenarios. Current explanation datasets often employ…