73 citations · 77 across the 7 of their papers we have counts for
7 papers · 1 filter
Optimistic Model Rollouts for Pessimistic Offline Policy Optimization
Yuanzhao Zhai, Yiying Li, Zijian Gao +5
Model-based offline reinforcement learning (RL) has made remarkable progress, offering a promising avenue for improving generalization with synthetic model rollouts. Existing works…
Nuclear Norm Maximization Based Curiosity-Driven Learning
Chao Chen, Zijian Gao, Kele Xu +5
To handle the sparsity of the extrinsic rewards in reinforcement learning, researchers have proposed intrinsic reward which enables the agent to learn the skills that might come in…
A Fixed Version of Quadratic Program in Gradient Episodic Memory
Wei Zhou, Yiying Li
Gradient Episodic Memory is indeed a novel method for continual learning, which solves new problems quickly without forgetting previously acquired knowledge. However, in the proces…
FedH2L: Federated Learning with Model and Statistical Heterogeneity
Yiying Li, Wei Zhou, Huaimin Wang +2
Federated learning (FL) enables distributed participants to collectively learn a strong global model without sacrificing their individual data privacy. Mainstream FL approaches req…
Online Meta-Critic Learning for Off-Policy Actor-Critic Methods
Wei Zhou, Yiying Li, Yongxin Yang +2
Off-Policy Actor-Critic (Off-PAC) methods have proven successful in a variety of continuous control tasks. Normally, the critic's action-value function is updated using temporal-di…
Feature Fusion Detector for Semantic Cognition of Remote Sensing
Wei Zhou, Yiying Li
The value of remote sensing images is of vital importance in many areas and needs to be refined by some cognitive approaches. The remote sensing detection is an appropriate way to…