activity
20232026
collaborators

9 papers

cs.LG2026

Group-Graph Policy Optimization for Long-Horizon Agentic Reinforcement Learning

Yunan Wang, Minghui Song, Zihan Zhang +6

Group-based Reinforcement Learning (RL) has significantly enhanced Large Language Models (LLMs) in agentic scenarios. To achieve finer-grained policy updates, recent agentic RL fra…

math.OC2026

Reachability-Augmented Dual Dynamic Programming for Optimal Path Parameterization

Yunan Wang, Jizhou Yan, Chuxiong Hu +1

Optimal path parameterization (OPP) is a fundamental problem for planning trajectories along a prescribed geometric path under kinodynamic constraints and task-dependent objectives…

math.OC2026

Time-Optimal Switching Surfaces for Triple Integrator under Full Box Constraints

Yunan Wang, Chuxiong Hu, Zhao Jin

Time-optimal control for triple integrator under full box constraints is a fundamental problem in the field of optimal control, which has been widely applied in the industry. Howev…

math.OC2024

A Novel State-Centric Necessary Condition for Time-Optimal Control of Controllable Linear Systems Based on Augmented Switching Laws (Extended Version)

Yunan Wang, Chuxiong Hu, Yujie Lin +3

Most existing necessary conditions for optimal control based on adjoining methods require both state and costate information, yet the unobservability of costates for a given feasib…

math.OC2024

Chattering Phenomena in Time-Optimal Control for High-Order Chain-of-Integrator Systems with Full State Constraints (Extended Version)

Yunan Wang, Chuxiong Hu, Zeyang Li +3

Time-optimal control for high-order chain-of-integrator systems with full state constraints remains an open and challenging problem within the discipline of optimal control. The be…

eess.SY2023

Time-Optimal Control for High-Order Chain-of-Integrators Systems with Full State Constraints and Arbitrary Terminal States (Extended Version)

Yunan Wang, Chuxiong Hu, Zeyang Li +3

Time-optimal control for high-order chain-of-integrators systems with full state constraints and arbitrarily given terminal states remains a challenging problem in the optimal cont…