activity
20242026
collaborators

11 papers

cs.LG2026

Fast A/B/n Testing: Exact Multi-Policy Comparison via Tree-Coupled Feedback Sharing

Yuxiao Wen

Online platforms increasingly compare many adaptive decision policies---ranking systems, recommendation algorithms, pricing rules, and language-model agents---while each reward-bea…

cs.LG2026

Learning to Bid with Unknown Private Values in Budget-Constrained First-Price Auctions

Zihao Hu, Yuxiao Wen, Yuan Yao +2

We study the operational problem of automated bidding in repeated first-price auctions under budget and return-on-spend (RoS) constraints. In this setting, an auto-bidder must tran…

cs.GT2026

The (Marginal) Value of a Search Ad: An Online Causal Framework for Repeated Second-price Auctions

Yuxiao Wen, Zihao Hu, Yanjun Han +2

Existing auto-bidding algorithms in digital advertising often treat the value of an ad opportunity as the revenue obtained when an ad is shown and/or clicked, and bid accordingly.…

cs.LG2026

Joint Value Estimation and Bidding in Repeated First-Price Auctions

Yuxiao Wen, Yanjun Han, Zhengyuan Zhou

We study regret minimization in repeated first-price auctions (FPAs), where a bidder observes only the realized outcome after each auction -- win or loss. This setup reflects pract…

cs.LG2025

Optimal Arm Elimination Algorithms for Combinatorial Bandits

Yuxiao Wen, Yanjun Han, Zhengyuan Zhou

Combinatorial bandits extend the classical bandit framework to settings where the learner selects multiple arms in each round, motivated by applications such as online recommendati…

cs.HC2025

ClassMind: Scaling Classroom Observation and Instructional Feedback with Multimodal AI

Ao Qu, Yuxi Wen, Jiayi Zhang +6

Classroom observation -- one of the most effective methods for teacher development -- remains limited due to high costs and a shortage of expert coaches. We present ClassMind, an A…