3 papers
cs.LG2026
Inverse RL Helps Align AI by Imitating Humans
Michał Wiliński, Liu Leqi, Chirag Nagpal
Language model alignment aims to make model behavior reliably reflect desirable properties such as helpfulness, safety, and instruction following. Current approaches typically use…
cs.LG2024
A Unified Causal Framework for Auditing Recommender Systems for Ethical Concerns
Vibhhu Sharma, Shantanu Gupta, Nil-Jana Akpinar +2
As recommender systems become widely deployed in different domains, they increasingly influence their users' beliefs and preferences. Auditing recommender systems is crucial as it…
cs.LG2018
Sample Complexity of Nonparametric Semi-Supervised Learning
Chen Dan, Liu Leqi, Bryon Aragam +2
We study the sample complexity of semi-supervised learning (SSL) and introduce new assumptions based on the mismatch between a mixture model learned from unlabeled data and the tru…