6 papers
MultiView-Bench: A Diagnostic Benchmark for World-Centric Multi-View Integration in VLMs
Hantao Zhang, Jinru Sui, Ed Li +2
Recent benchmarks for VLMs largely assess single- or limited-view perception, leaving untested the core cognitive ability to integrate observations across viewpoints into a coheren…
The Future of Artificial Intelligence and the Mathematical and Physical Sciences (AI+MPS)
Andrew Ferguson, Marisa LaFleur, Lars Ruthotto +97
This community paper developed out of the NSF Workshop on the Future of Artificial Intelligence (AI) and the Mathematical and Physics Sciences (MPS), which was held in March 2025 w…
Learning Task Representations from In-Context Learning
Baturay Saglam, Xinyang Hu, Zhuoran Yang +2
Large language models (LLMs) have demonstrated remarkable proficiency in in-context learning (ICL), where models adapt to new tasks through example-based prompts without requiring…
Build Your Personalized Research Group: A Multiagent Framework for Continual and Interactive Science Automation
Ed Li, Junyu Ren, Xintian Pan +4
The automation of scientific discovery represents a critical milestone in Artificial Intelligence (AI) research. However, existing agentic systems for science suffer from two funda…
Learning to Lead: Incentivizing Strategic Agents in the Dark
Yuchen Wu, Xinyi Zhong, Zhuoran Yang
We study an online learning version of the generalized principal-agent model, where a principal interacts repeatedly with a strategic agent possessing private types, private reward…
Physical Informed Driving World Model
Zhuoran Yang, Xi Guo, Chenjing Ding +2
Autonomous driving requires robust perception models trained on high-quality, large-scale multi-view driving videos for tasks like 3D object detection, segmentation and trajectory…