13 citations · 15 across the 4 of their papers we have counts for
4 papers
CLARIFY: Contrastive Preference Reinforcement Learning for Untangling Ambiguous Queries
Ni Mu, Hao Hu, Xiao Hu +3
Preference-based reinforcement learning (PbRL) bypasses explicit reward engineering by inferring reward functions from human preference comparisons, enabling better alignment with…
Mind the Gap: Offline Policy Optimization for Imperfect Rewards
Jianxiong Li, Xiao Hu, Haoran Xu +4
Reward function is essential in reinforcement learning (RL), serving as the guiding signal to incentivize agents to solve given tasks, however, is also notoriously difficult to des…
A Survey of ADMM Variants for Distributed Optimization: Problems, Algorithms and Features
Yu Yang, Xiaohong Guan, Qing-Shan Jia +3
By coordinating terminal smart devices or microprocessors to engage in cooperative computation to achieve systemlevel targets, distributed optimization is incrementally favored by…
A Two-phase On-line Joint Scheduling for Welfare Maximization of Charging Station
Qilong Huang, Qing-Shan Jia, Xiang Wu +2
The large adoption of EVs brings practical interest to the operation optimization of the charging station. The joint scheduling of pricing and charging control will achieve a win-w…