4 papers
A Unified Framework for Scalable and Robust Paper Assignment
Michael Cui, Chenxin Dai, Yixuan Even Xu +1
Assigning papers to reviewers is a central challenge in the peer-review process of large academic conferences. Program chairs must balance competing objectives, including maximizin…
Not All Rollouts are Useful: Down-Sampling Rollouts in LLM Reinforcement Learning
Yixuan Even Xu, Yash Savani, Fei Fang +1
Reinforcement learning with verifiable rewards (RLVR) has emerged as the leading approach for enhancing reasoning capabilities in large language models. However, it faces a fundame…
Improving Community-Participated Patrol for Anti-Poaching
Yufei Wu, Yixuan Even Xu, Xuming Zhang +3
Community engagement plays a critical role in anti-poaching efforts, yet existing mathematical models aimed at enhancing this engagement often overlook direct participation by comm…
Deviate or Not: Learning Coalition Structures with Multiple-bit Observations in Games
Yixuan Even Xu, Zhe Feng, Fei Fang
We consider the Coalition Structure Learning (CSL) problem in multi-agent systems, motivated by the existence of coalitions in many real-world systems, e.g., trading platforms and…