◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yu Liu

4 papers hereh-index 562 citations12 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author1
  • last author2

Across the 3 of 4 papers where every author was matched, so the position is known.

fields
  • cs.LG4
same name
  • Yu Liu — 49 papers, h 21
  • Yu Liu — 45 papers, h 32
  • Yu Liu — 22 papers, h 18
  • Yu Liu — 21 papers, h 11
  • Yu Liu — 16 papers, h 14
  • Yu Liu — 15 papers, h 21

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedMEPG: A Minimalist Ensemble Policy Gradient Framework for Deep Reinforcement Learning

6 citations · 7 across the 4 of their papers we have counts for

collaborators

4 papers

cs.LG2023

Centralized Cooperative Exploration Policy for Continuous Control Tasks

Chao Li, Chen Gong, Qiang He +2

The deep reinforcement learning (DRL) algorithm works brilliantly on solving various complex control tasks. This phenomenal success can be partly attributed to DRL encouraging inte…

cs.LG2021

The f-Divergence Reinforcement Learning Framework

Chen Gong, Qiang He, Yunpeng Bai +6

The framework of deep reinforcement learning (DRL) provides a powerful and widely applicable mathematical formalization for sequential decision-making. This paper present a novel D…

cs.LG2021★ 1 cited

LDC-VAE: A Latent Distribution Consistency Approach to Variational AutoEncoders

Xiaoyu Chen, Chen Gong, Qiang He +2

Variational autoencoders (VAEs), as an important aspect of generative models, have received a lot of research interests and reached many successful applications. However, it is alw…

cs.LG2021★ 6 cited

MEPG: A Minimalist Ensemble Policy Gradient Framework for Deep Reinforcement Learning

Qiang He, Huangyuan Su, Chen Gong +1

During the training of a reinforcement learning (RL) agent, the distribution of training data is non-stationary as the agent's behavior changes over time. Therefore, there is a ris…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.