8 citations · 21 across the 10 of their papers we have counts for
4 papers · 1 filter
On the Convergence Theory of Meta Reinforcement Learning with Personalized Policies
Haozhi Wang, Qing Wang, Yunfeng Shao +3
Modern meta-reinforcement learning (Meta-RL) methods are mainly developed based on model-agnostic meta-learning, which performs policy gradient steps across tasks to maximize polic…
GALOIS: Boosting Deep Reinforcement Learning via Generalizable Logic Synthesis
Yushi Cao, Zhiming Li, Tianpei Yang +5
Despite achieving superior performance in human-level control problems, unlike humans, deep reinforcement learning (DRL) lacks high-order intelligence (e.g., logic deduction and re…
Revisiting QMIX: Discriminative Credit Assignment by Gradient Entropy Regularization
Jian Zhao, Yue Zhang, Xunhan Hu +5
In cooperative multi-agent systems, agents jointly take actions and receive a team reward instead of individual rewards. In the absence of individual reward signals, credit assignm…
Ranking Cost: Building An Efficient and Scalable Circuit Routing Planner with Evolution-Based Optimization
Shiyu Huang, Bin Wang, Dong Li +3
Circuit routing has been a historically challenging problem in designing electronic systems such as very large-scale integration (VLSI) and printed circuit boards (PCBs). The main…