11 papers
Fast A/B/n Testing: Exact Multi-Policy Comparison via Tree-Coupled Feedback Sharing
Yuxiao Wen
Online platforms increasingly compare many adaptive decision policies---ranking systems, recommendation algorithms, pricing rules, and language-model agents---while each reward-bea…
Learning to Bid with Unknown Private Values in Budget-Constrained First-Price Auctions
Zihao Hu, Yuxiao Wen, Yuan Yao +2
We study the operational problem of automated bidding in repeated first-price auctions under budget and return-on-spend (RoS) constraints. In this setting, an auto-bidder must tran…
The (Marginal) Value of a Search Ad: An Online Causal Framework for Repeated Second-price Auctions
Yuxiao Wen, Zihao Hu, Yanjun Han +2
Existing auto-bidding algorithms in digital advertising often treat the value of an ad opportunity as the revenue obtained when an ad is shown and/or clicked, and bid accordingly.…
Joint Value Estimation and Bidding in Repeated First-Price Auctions
Yuxiao Wen, Yanjun Han, Zhengyuan Zhou
We study regret minimization in repeated first-price auctions (FPAs), where a bidder observes only the realized outcome after each auction -- win or loss. This setup reflects pract…
Optimal Arm Elimination Algorithms for Combinatorial Bandits
Yuxiao Wen, Yanjun Han, Zhengyuan Zhou
Combinatorial bandits extend the classical bandit framework to settings where the learner selects multiple arms in each round, motivated by applications such as online recommendati…
ClassMind: Scaling Classroom Observation and Instructional Feedback with Multimodal AI
Ao Qu, Yuxi Wen, Jiayi Zhang +6
Classroom observation -- one of the most effective methods for teacher development -- remains limited due to high costs and a shortage of expert coaches. We present ClassMind, an A…