33 citations · 122 across the 37 of their papers we have counts for
Showing 2023 · cs.LGShow all
2 papers · 2 filters
cs.LG2023
Efficiently Escaping Saddle Points for Policy Optimization
Sadegh Khorasani, Saber Salehkaleybar, Negar Kiyavash +2
Policy gradient (PG) is widely used in reinforcement learning due to its scalability and good performance. In recent years, several variance-reduced PG methods have been proposed w…
cs.LG2023
Causal Imitability Under Context-Specific Independence Relations
Fateme Jamshidi, Sina Akbari, Negar Kiyavash
Drawbacks of ignoring the causal mechanisms when performing imitation learning have recently been acknowledged. Several approaches both to assess the feasibility of imitation and t…