1 paper
Mehrdad Mohammadi, Qi Zheng, Ruoqing Zhu
We propose an (offline) multi-dimensional distributional reinforcement learning framework (KE-DRL) that leverages Hilbert space mappings to estimate the kernel mean embedding of th…