5 papers · 1 filter
Architectural Implications of Agentic AI Workflows
Jirong Yang, Peizhe Liu, Chaojie Zhang +1
Agentic AI is emerging in datacenters, but its architectural implications remain unexplored. We organize agentic workflows in a taxonomy and present its first architectural charact…
LiveOIBench: Can Large Language Models Outperform Human Contestants in Informatics Olympiads?
Kaijian Zou, Aaron Xiong, Yunxiang Zhang +6
Competitive programming problems are increasingly used to evaluate the coding capabilities of large language models (LLMs) due to their complexity and ease of verification. Yet, cu…
Agents' Last Exam
Yiyou Sun, Xinyang Han, Weichen Zhang +306
Recent AI systems have achieved strong results on a wide range of benchmarks, yet these gains have not translated into economically meaningful deployment across many professional d…
Aligning LLM agents with human learning and adjustment behavior: a dual agent approach
Tianming Liu, Jirong Yang, Yafeng Yin +3
Effective modeling of how human travelers learn and adjust their travel behavior from interacting with transportation systems is critical for system assessment and planning. Howeve…
Toward LLM-Agent-Based Modeling of Transportation Systems: A Conceptual Framework
Tianming Liu, Jirong Yang, Yafeng Yin
In transportation system demand modeling and simulation, agent-based models and microsimulations are current state-of-the-art approaches. However, existing agent-based models still…