Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Closing the Intent-to-Behavior Gap via Fulfillment Priority Logic
Bassel El Mabsout, Abdelrahman Abdelgawad, Renato Mancuso
Practitioners designing reinforcement learning policies face a fundamental challenge: translating intended behavioral objectives into representative reward functions. This challeng…
cs.LG2021
Honey, I Shrunk The Actor: A Case Study on Preserving Performance with Smaller Actors in Actor-Critic RL
Siddharth Mysore, Bassel Mabsout, Renato Mancuso +1
Actors and critics in actor-critic reinforcement learning algorithms are functionally separate, yet they often use the same network architectures. This case study explores the perf…