1 citations · 1 across the 5 of their papers we have counts for
3 papers · 1 filter
Locally Adaptive Multi-Objective Learning
Jivat Neet Kaur, Isaac Gibbs, Michael I. Jordan
We consider the general problem of learning a predictor that satisfies multiple objectives of interest simultaneously, a broad framework that captures a range of specific learning…
How Sampling Shapes LLM Alignment: From One-Shot Optima to Iterative Dynamics
Yurong Chen, Yu He, Michael I. Jordan +1
Standard methods for aligning large language models with human preferences learn from pairwise comparisons among sampled candidate responses and regularize toward a reference polic…
Reduced-Rank Multi-objective Policy Learning and Optimization
Ezinne Nwankwo, Michael I. Jordan, Angela Zhou
Evaluating the causal impacts of possible interventions is crucial for informing decision-making, especially towards improving access to opportunity. However, if causal effects are…