2 papers
cs.AI2026
Best Arm Identification in Generalized Linear Bandits via Hybrid Feedback
Qirun Zeng, Xuchuang Wang, Jiayi Shen +3
We study fixed-confidence best arm identification in generalized linear bandits under a hybrid feedback model: at each round, the learner may query either (i) absolute reward feedb…
cs.LG2024
GO4Align: Group Optimization for Multi-Task Alignment
Jiayi Shen, Cheems Wang, Zehao Xiao +2
This paper proposes \textit{GO4Align}, a multi-task optimization approach that tackles task imbalance by explicitly aligning the optimization across tasks. To achieve this, we desi…