low-resource languages 1meta-learning 1multilingual alignment 1preference learning 1reinforcement learning from human feedback 1
From the 1 of 20 linked papers with an AI index.
Showing stat.MLShow all
2 papers · 1 filter
stat.ML2026
Sparse Offline Reinforcement Learning with Corruption Robustness
Nam Phuong Tran, Andi Nika, Goran Radanovic +2
We investigate robustness to strong data corruption in offline sparse reinforcement learning (RL). In our setting, an adversary may arbitrarily perturb a fraction of the collected…
stat.ML2025
Symmetric Linear Bandits with Hidden Symmetry
Nam Phuong Tran, The Anh Ta, Debmalya Mandal +1
High-dimensional linear bandits with low-dimensional structure have received considerable attention in recent studies due to their practical significance. The most common structure…