From the 1 of 4 linked papers with an AI index.
4 papers
Meta-Learning Preferences for Multilingual LLM Alignment
Jiaying Lin, Seongho Son, Nam Phuong Tran +3
The paper introduces a meta-learning method that uses preference data from high-resource languages to quickly adapt large language models to low-resource languages with very few hu…
Sparse Offline Reinforcement Learning with Corruption Robustness
Nam Phuong Tran, Andi Nika, Goran Radanovic +2
We investigate robustness to strong data corruption in offline sparse reinforcement learning (RL). In our setting, an adversary may arbitrarily perturb a fraction of the collected…
Symmetric Linear Bandits with Hidden Symmetry
Nam Phuong Tran, The Anh Ta, Debmalya Mandal +1
High-dimensional linear bandits with low-dimensional structure have received considerable attention in recent studies due to their practical significance. The most common structure…
Learning the Expected Core of Strictly Convex Stochastic Cooperative Games
Nam Phuong Tran, The Anh Ta, Shuqing Shi +3
Reward allocation, also known as the credit assignment problem, has been an important topic in economics, engineering, and machine learning. An important concept in reward allocati…