3 papers
cs.AI2026
LitReview Arena: Evaluating Literature Review Agents with Battle-Style Peer Review Platform
Ruotong Zhao, Zhiyu Chen, Xurui Liu +7
Literature reviews are essential to scientific progress, but rigorously evaluating automatically generated reviews remains difficult because many aspects of research utility depend…
cs.CY2025
OmniScientist: Toward a Co-evolving Ecosystem of Human and AI Scientists
Chenyang Shao, Dehao Huang, Yu Li +18
With the rapid development of Large Language Models (LLMs), AI agents have demonstrated increasing proficiency in scientific tasks, ranging from hypothesis generation and experimen…
cs.AI2025
CrimeMind: Simulating Urban Crime with Multi-Modal LLM Agents
Qingbin Zeng, Ruotong Zhao, Jinzhu Mao +3
Modeling urban crime is an important yet challenging task that requires understanding the subtle visual, social, and cultural cues embedded in urban environments. Previous work has…