◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Xiao Hu

2 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.LG2
ORCID 0000-0003-3994-0385
same name
  • Xiao Hu — 34 papers, h 33
  • Xiao Hu — 7 papers
  • Xiao Hu — 5 papers
  • Xiao Hu — 4 papers, h 6
  • Xiao Hu — 3 papers
  • Xiao Hu — 2 papers, h 8

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedPROTO: Iterative Policy Regularized Offline-to-Online Reinforcement Learning

5 citations · 6 across the 2 of their papers we have counts for

collaborators

2 papers

cs.LG2023★ 5 cited

PROTO: Iterative Policy Regularized Offline-to-Online Reinforcement Learning

Jianxiong Li, Xiao Hu, Haoran Xu +3

Offline-to-online reinforcement learning (RL), by combining the benefits of offline pretraining and online finetuning, promises enhanced sample efficiency and policy performance. H…

cs.LG2023★ 1 cited

Mind the Gap: Offline Policy Optimization for Imperfect Rewards

Jianxiong Li, Xiao Hu, Haoran Xu +4

Reward function is essential in reinforcement learning (RL), serving as the guiding signal to incentivize agents to solve given tasks, however, is also notoriously difficult to des…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.