15 citations · 15 across the 3 of their papers we have counts for
3 papers
cs.LG2024
Scale-free Adversarial Reinforcement Learning
Mingyu Chen, Xuezhou Zhang
This paper initiates the study of scale-free learning in Markov Decision Processes (MDPs), where the scale of rewards/losses is unknown to the learner. We design a generic algorith…
stat.ML2023
Improved Algorithms for Adversarial Bandits with Unbounded Losses
Mingyu Chen, Xuezhou Zhang
We consider the Adversarial Multi-Armed Bandits (MAB) problem with unbounded losses, where the algorithms have no prior knowledge on the sizes of the losses. We present UMAB-NN and…
cs.CL2023★ 15 cited
Chain-of-Thought Hub: A Continuous Effort to Measure Large Language Models' Reasoning Performance
Yao Fu, Litu Ou, Mingyu Chen +3
As large language models (LLMs) are continuously being developed, their evaluation becomes increasingly important yet challenging. This work proposes Chain-of-Thought Hub, an open-…