6 papers
World Models in Pieces: Structural Certification for General Agents
Yikai Lu, Yifei Wu, Xinyu Lu +1
In the big-world regime, agents cannot be universally capable and their ability is inevitably specialized across a world model in pieces. Consequently, standard uniform guarantees…
EComAgentBench: Benchmarking Shopping Agents on Long-Horizon Tasks with Distributed Hidden Intent
Zeyao Du, Tong Li, Yanci Zhang +1
As LLM-based shopping agents enter production, existing benchmarks fail to capture how a shopper's requirements arrive: stated implicitly in the query, recorded in a profile, or re…
Bridging Natural Language and Microgrid Dynamics: A Context-Aware Simulator and Dataset
Tinko Sebastian Bartels, Ruixiang Wu, Xinyu Lu +5
Addressing the critical need for intelligent, context-aware energy management in renewable systems, we introduce the OpenCEM Simulator and Dataset: the first open-source digital tw…
DCT-MARL: A Dynamic Communication Topology-Based MARL Algorithm for Connected Vehicle Platoon Control
Yaqi Xu, Yan Shi, Jin Tian +4
With the rapid advancement of vehicular communication facilities and autonomous driving technologies, connected vehicle platooning has emerged as a promising approach to improve tr…
Beyond Numeric Rewards: In-Context Dueling Bandits with LLM Agents
Fanzeng Xia, Hao Liu, Yisong Yue +1
In-Context Reinforcement Learning (ICRL) is a frontier paradigm to solve Reinforcement Learning (RL) problems in the foundation model era. While ICRL capabilities have been demonst…
Rethinking the Unsolvable: When In-Context Search Meets Test-Time Scaling
Fanzeng Xia, Yidong Luo, Tinko Sebastian Bartels +2
Recent research has highlighted that Large Language Models (LLMs), even when trained to generate extended long reasoning steps, still face significant challenges on hard reasoning…