13 citations · 19 across the 2 of their papers we have counts for
3 papers
cs.LG2025
Automatic Reward Shaping from Multi-Objective Human Heuristics
Yuqing Xie, Jiayu Chen, Wenhao Tang +3
Designing effective reward functions remains a central challenge in reinforcement learning, especially in multi-objective environments. In this work, we propose Multi-Objective Rew…
cs.AI2021★ 13 cited
Discovering Diverse Multi-Agent Strategic Behavior via Reward Randomization
Zhenggang Tang, Chao Yu, Boyuan Chen +6
We propose a simple, general and effective technique, Reward Randomization for discovering diverse strategic policies in complex multi-agent games. Combining reward randomization a…
cs.DB2020★ 6 cited
SQLFlow: A Bridge between SQL and Machine Learning
Yi Wang, Yang Yang, Weiguo Zhu +11
Industrial AI systems are mostly end-to-end machine learning (ML) workflows. A typical recommendation or business intelligence system includes many online micro-services and offlin…