16 citations · 16 across the 1 of their papers we have counts for
1 paper
Pavithra Harsha, Ashish Jagmohan, Jayant Kalagnanam +2
We present a Reinforcement Learning (RL) based framework for optimizing long-term discounted reward problems with large combinatorial action space and state dependent constraints.…