activity
20232026
most citedSkillTree: Explainable Skill-Based Deep Reinforcement Learning for Long-Horizon Control Tasks

4 citations · 4 across the 11 of their papers we have counts for

collaborators
Showing cs.LGShow all

6 papers · 1 filter

cs.LG2026

Curriculum reinforcement learning with measurable task representation learning

Yongyan Wen, Siyuan Li, Mingjian Fu +3

In curriculum reinforcement learning (CRL), an agent incrementally accumulates knowledge over a sequence of tasks (i.e., a curriculum), and the learning process is aimed at using t…

cs.LG2025

Unsupervised Skill Discovery through Skill Regions Differentiation

Ting Xiao, Jiakun Zheng, Rushuai Yang +4

Unsupervised Reinforcement Learning (RL) aims to discover diverse behaviors that can accelerate the learning of downstream tasks. Previous methods typically focus on entropy-based…

cs.LG2024

SkillTree: Explainable Skill-Based Deep Reinforcement Learning for Long-Horizon Control Tasks

Yongyan Wen, Siyuan Li, Rongchang Zuo +3

Deep reinforcement learning (DRL) has achieved remarkable success in various research domains. However, its reliance on neural networks results in a lack of transparency, which lim…

cs.LG2024

An Imitative Reinforcement Learning Framework for Pursuit-Lock-Launch Missions

Siyuan Li, Rongchang Zuo, Bofei Liu +3

Unmanned Combat Aerial Vehicle (UCAV) Within-Visual-Range (WVR) engagement, referring to a fight between two or more UCAVs at close quarters, plays a decisive role on the aerial ba…

cs.LG2024

Auxiliary Reward Generation with Transition Distance Representation Learning

Siyuan Li, Shijie Han, Yingnan Zhao +2

Reinforcement learning (RL) has shown its strength in challenging sequential decision-making problems. The reward function in RL is crucial to the learning performance, as it serve…

cs.LG2023

Robust Visual Imitation Learning with Inverse Dynamics Representations

Siyuan Li, Xun Wang, Rongchang Zuo +5

Imitation learning (IL) has achieved considerable success in solving complex sequential decision-making problems. However, current IL methods mainly assume that the environment for…