2 citations · 3 across the 5 of their papers we have counts for
Showing cs.ROShow all
2 papers · 1 filter
cs.RO2024★ 18 cited
Theory of Mind abilities of Large Language Models in Human-Robot Interaction : An Illusion?
Mudit Verma, Siddhant Bhambri, Subbarao Kambhampati
Large Language Models have shown exceptional generative abilities in various natural language and generation tasks. However, possible anthropomorphization and leniency towards fail…
cs.RO2023★ 2 cited
Exploiting Unlabeled Data for Feedback Efficient Human Preference based Reinforcement Learning
Mudit Verma, Siddhant Bhambri, Subbarao Kambhampati
Preference Based Reinforcement Learning has shown much promise for utilizing human binary feedback on queried trajectory pairs to recover the underlying reward model of the Human i…