2 papers
cs.CL2026
Structured In-context Environment Scaling for Large Language Model Reasoning
Peng Yu, Zeyuan Zhao, Shao Zhang +3
Large language models (LLMs) have achieved significant advancements in reasoning capabilities through reinforcement learning (RL) via environmental exploration. As the intrinsic pr…
cs.RO2025
Sequence Pathfinder for Multi-Agent Pickup and Delivery in the Warehouse
Zeyuan Zhao, Chaoran Li, Shao Zhang +1
Multi-Agent Pickup and Delivery (MAPD) is a challenging extension of Multi-Agent Path Finding (MAPF), where agents are required to sequentially complete tasks with fixed-location p…