1 paper · 1 filter
Ayushman Singh, Siddharth Aphale
The paper investigates why bilinear contrastive critics, which rank actions for reinforcement learning policies, can produce unsafe or misleading rankings due to issues like norm d…