1 paper
Wenhua Nie, Binhan Luo, Zijie Meng +2
Large language models can score well on named game-theory benchmarks while failing on the same strategic computation once semantic cues are removed. We show this gap with procedura…