3 papers
cs.LG2026
IV-ICL: Bounding Causal Effects with Instrumental Variables via In-Context Learning
Vahid Balazadeh, Hamidreza Kamkari, Medha Barath +2
The instrumental-variables (IV) setting is standard for partial identification of causal effects when unobserved confounding makes point identification impossible. Existing approac…
cs.LG2024
Personalized Adaptation via In-Context Preference Learning
Allison Lau, Younwoo Choi, Vahid Balazadeh +3
Reinforcement Learning from Human Feedback (RLHF) is widely used to align Language Models (LMs) with human preferences. However, existing approaches often neglect individual user p…
cs.LG2024
Sequential Decision Making with Expert Demonstrations under Unobserved Heterogeneity
Vahid Balazadeh, Keertana Chidambaram, Viet Nguyen +2
We study the problem of online sequential decision-making given auxiliary demonstrations from experts who made their decisions based on unobserved contextual information. These dem…