works on

From the 1 of 5 linked papers with an AI index.

activity
20242026
collaborators

5 papers

stat.ML2026

Escaping Model Collapse via Synthetic Data Verification: Near-term Improvements and Long-term Convergence

Bingji Yi, Qiyuan Liu, Yuwei Cheng +1

The paper proposes using an external synthetic data verifier to prevent model collapse during iterative retraining of generative models on their own synthetic data, providing theor…

cs.CL2026

Fine-Tuning Improves Information Conveyance in Language Models

Yuwei Cheng, Weiyi Tian, Haifeng Xu

Fine-tuning is often believed to reduce uncertainty and diversity in large language models, but existing analyses overlook output length, a key confounder, and therefore fail to ca…

cs.LG2025

Learning Personalized Ad Impact via Contextual Reinforcement Learning under Delayed Rewards

Yuwei Cheng, Zifeng Zhao, Haifeng Xu

Online advertising platforms use automated auctions to connect advertisers with potential customers, requiring effective bidding strategies to maximize profits. Accurate ad impact…

cs.HC2025

Can LLMs Address Mental Health Questions? A Comparison with Human Therapists

Synthia Wang, Yuwei Cheng, Austin Song +3

Limited access to mental health care has motivated the use of digital tools and conversational agents powered by large language models (LLMs), yet their quality and reception remai…

cs.LG2024

Learning from Imperfect Human Feedback: a Tale from Corruption-Robust Dueling

Yuwei Cheng, Fan Yao, Xuefeng Liu +1

This paper studies Learning from Imperfect Human Feedback (LIHF), addressing the potential irrationality or imperfect perception when learning from comparative human feedback. Buil…