2 papers
cs.LG2023
Policy composition in reinforcement learning via multi-objective policy optimization
Shruti Mishra, Ankit Anand, Jordan Hoffmann +4
We enable reinforcement learning agents to learn successful behavior policies by utilizing relevant pre-existing teacher policies. The teacher policies are introduced as objectives…
hep-th2023
Island in Warped AdS Black Holes
Ankit Anand
This paper investigates the Page curve in Warped Anti-de Sitter black holes using the quantum extremal surface prescription. The findings reveal that in the absence of an island, t…