8 citations · 10 across the 4 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2022★ 1 cited
Symbol Guided Hindsight Priors for Reward Learning from Human Preferences
Mudit Verma, Katherine Metcalf
Specifying rewards for reinforcement learned (RL) agents is challenging. Preference-based RL (PbRL) mitigates these challenges by inferring a reward from feedback over sets of traj…
cs.LG2021★ 8 cited
A Novel Framework for Neural Architecture Search in the Hill Climbing Domain
Mudit Verma, Pradyumna Sinha, Karan Goyal +2
Neural networks have now long been used for solving complex problems of image domain, yet designing the same needs manual expertise. Furthermore, techniques for automatically gener…