2 papers
cs.LG2026
Towards Sustainable Investment Policies Informed by Opponent Shaping
Juan Agustin Duque, Razvan Ciuca, Ayoub Echchahed +2
Addressing climate change requires global coordination, yet rational economic actors often prioritize immediate gains over collective welfare, resulting in social dilemmas. InvestE…
cs.LG2025
Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL
Ghada Sokar, Johan Obando-Ceron, Aaron Courville +2
The use of deep neural networks in reinforcement learning (RL) often suffers from performance degradation as model size increases. While soft mixtures of experts (SoftMoEs) have re…