2 citations · 2 across the 8 of their papers we have counts for
Showing 2026Show all
3 papers · 1 filter
cs.LG2026
Bayesian Best-Arm Identification with Abstention: A Polynomial-to-Exponential Phase Transition
Yuqi Huang, Yunlong Hou, Vincent Y. F. Tan
We study the Bayesian fixed-budget best-arm identification problem in which a learner can abstain from making a terminal recommendation. Subject to an abstention budget , we ana…
cs.LG2026
On the Benefits of Free Exploration for Regret Minimization in Multi-Armed Bandits
Yunlong Hou, Zixin Zhong, Vincent Y. F. Tan
We study a stochastic multi-armed bandit problem where an agent is granted a free exploration budget before regret accumulates, a setting not captured by the classic regret minimiz…
cs.LG2026
Demystifying the Slash Pattern in Attention: The Role of RoPE
Yuan Cheng, Fengzhuo Zhang, Yunlong Hou +5
Large Language Models (LLMs) often exhibit slash attention patterns, where attention scores concentrate along the -th sub-diagonal for some offset . These patterns play a key…