1 citations · 1 across the 7 of their papers we have counts for
Showing 2025Show all
2 papers · 1 filter
stat.ML2025
Uncertainty Quantification for Large Language Model Reward Learning under Heterogeneous Human Feedback
Pangpang Liu, Junwei Lu, Will Wei Sun
We study estimation and statistical inference for reward models used in aligning large language models (LLMs). A key component of LLM alignment is reinforcement learning from human…
cs.GT2025
Fairness-aware Contextual Dynamic Pricing with Strategic Buyers
Pangpang Liu, Will Wei Sun
Contextual pricing strategies are prevalent in online retailing, where the seller adjusts prices based on products' attributes and buyers' characteristics. Although such strategies…