4 citations · 7 across the 51 of their papers we have counts for
3 papers · 2 filters
Beyond Policy Optimization: A Data Curation Flywheel for Sparse-Reward Long-Horizon Planning
Yutong Wang, Pengliang Ji, Kaixin Li +3
Large Language Reasoning Models have demonstrated remarkable success on static tasks, yet their application to multi-round agentic planning in interactive environments faces two fu…
Multimodal Fused Learning for Solving the Generalized Traveling Salesman Problem in Robotic Task Planning
Jiaqi Cheng, Mingfeng Fan, Xuefeng Zhang +4
Effective and efficient task planning is essential for mobile robots, especially in applications like warehouse retrieval and environmental monitoring. These tasks often involve se…
Preference-Driven Multi-Objective Combinatorial Optimization with Conditional Computation
Mingfeng Fan, Jianan Zhou, Yifeng Zhang +3
Recent deep reinforcement learning methods have achieved remarkable success in solving multi-objective combinatorial optimization problems (MOCOPs) by decomposing them into multipl…