2 papers
cs.CV2026
VGI-Bench: Probing Visual Intelligence in Video Generation Models
Xuan He, Cong Wei, Yuhao Cheng +20
Recent studies suggest that video generation models can exhibit certain forms of zero-shot visual reasoning through generated frames. Yet reliable evaluation remains challenging: b…
cs.CL2026
Under the Influence: Quantifying Persuasion and Vigilance in Large Language Models
Sasha Robinson, Katherine M. Collins, Ilia Sucholutsky +1
With increasing integration of Large Language Models (LLMs) into areas of high-stakes human decision-making, it is important to understand the risks they introduce as advisors. To…