2 papers
cs.AI2026
Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games
Yidong He, Yutao Lai, Pengxu Yang +4
While Large Language Models (LLMs) excel in certain reasoning tasks, they struggle in multi-agent games where the final outcome depends on the joint strategies of all agents. In mu…
cs.CV2025
ViRectify: A Challenging Benchmark for Video Reasoning Correction with Multimodal Large Language Models
Xusen Hei, Jiali Chen, Jinyu Yang +2
As multimodal large language models (MLLMs) frequently exhibit errors in complex video reasoning scenarios, correcting these errors is critical for uncovering their weaknesses and…