2 papers
cs.AI2026
Agent-UCT: Upper Confidence Bounds Applied to Trees for Agentic Workflow Optimization with Cost-Awareness
Yang Li, Hai Liu, Dian Shao +8
Optimizing agentic workflows, such as retrieval-augmented generation (RAG) pipelines, requires navigating a combinatorial space of discrete component choices under tight evaluation…
cs.RO2026
MAPL: Multi-Objective Preference Learning for Robot Locomotion
Xiyue Chen, Muhan Lin, Shuyang Shi +1
Reward design remains a major bottleneck in reinforcement learning for robot locomotion, where successful policies often depend on carefully tuned, task-specific reward functions.…