Publications (8)
Representative Ranking for Deliberation in the Public Sphere
Manon Revel, Smitha Milli, Tyler Lu +2
Online comment sections, such as those on news sites or social media, have the potential to foster informal public deliberation, However, this potential is often undermined by the…
Aligning LLMs Toward Multi-Turn Conversational Outcomes Using Iterative PPO
Daniel R. Jiang, Jalaj Bhandari, Yukai Yang +2
Optimizing large language models (LLMs) for multi-turn conversational outcomes remains a significant challenge, especially in goal-oriented settings like AI marketing or sales agen…
Learning Low-Density Separators
Shai Ben-David, Tyler Lu, David Pal +1
We define a novel, basic, unsupervised learning problem - learning the lowest density homogeneous hyperplane separator of an unknown probability distribution. This task is relevant…
Safe Exploration for Identifying Linear Systems via Robust Optimization
Tyler Lu, Martin Zinkevich, Craig Boutilier +2
Safely exploring an unknown dynamical system is critical to the deployment of reinforcement learning (RL) in physical systems where failures may have catastrophic consequences. In…
Discovering Personalized Semantics for Soft Attributes in Recommender Systems using Concept Activation Vectors
Christina Göpfert, Alex Haig, Yinlam Chow +7
Interactive recommender systems have emerged as a promising paradigm to overcome the limitations of the primitive user feedback used by traditional recommender systems (e.g., click…
Bayesian Vote Manipulation: Optimal Strategies and Impact on Welfare
Tyler Lu, Pingzhong Tang, Ariel D. Procaccia +1
Most analyses of manipulation of voting schemes have adopted two assumptions that greatly diminish their practical import. First, it is usually assumed that the manipulators have f…