44 citations · 66 across the 10 of their papers we have counts for
3 papers · 1 filter
Guiding Reinforcement Learning Using Uncertainty-Aware Large Language Models
Maryam Shoaeinaeini, Brent Harrison
Human guidance in reinforcement learning (RL) is often impractical for large-scale applications due to high costs and time constraints. Large Language Models (LLMs) offer a promisi…
Using Non-Stationary Bandits for Learning in Repeated Cournot Games with Non-Stationary Demand
Kshitija Taywade, Brent Harrison, Judy Goldsmith
Many past attempts at modeling repeated Cournot games assume that demand is stationary. This does not align with real-world scenarios in which market demands can evolve over a prod…
Training Value-Aligned Reinforcement Learning Agents Using a Normative Prior
Md Sultan Al Nahian, Spencer Frazier, Brent Harrison +1
As more machine learning agents interact with humans, it is increasingly a prospect that an agent trained to perform a task optimally, using only a measure of task performance as f…