3 papers
cs.AI2026
Mitigating LLM biases toward spurious social contexts using direct preference optimization
Hyunji Nam, Dorottya Demszky
LLMs are increasingly used for high-stakes decision-making, yet their sensitivity to spurious contextual information can introduce harmful biases. This is a critical concern when m…
cs.AI2025
Predicting Long Term Sequential Policy Value Using Softer Surrogates
Hyunji Nam, Allen Nie, Ge Gao +2
Off-policy policy evaluation (OPE) estimates the outcome of a new policy using historical data collected from a different policy. However, existing OPE methods cannot handle cases…
cs.LG2024
Short-Long Policy Evaluation with Novel Actions
Hyunji Alex Nam, Yash Chandak, Emma Brunskill
From incorporating LLMs in education, to identifying new drugs and improving ways to charge batteries, innovators constantly try new strategies in search of better long-term outcom…