1 citations · 1 across the 2 of their papers we have counts for
3 papers
DeDPO: Debiased Direct Preference Optimization for Diffusion Models
Khiem Pham, Quang Nguyen, Tung Nguyen +4
Direct Preference Optimization (DPO) has emerged as a predominant alignment method for diffusion models, facilitating off-policy training without explicit reward modeling. However,…
Improving realistic semi-supervised learning with doubly robust estimation
Khiem Pham, Charles Herrmann, Ramin Zabih
A major challenge in Semi-Supervised Learning (SSL) is the limited information available about the class distribution in the unlabeled data. In many real-world applications this ar…
A Stable and Efficient Covariate-Balancing Estimator for Causal Survival Effects
Khiem Pham, David A. Hirshberg, Phuong-Mai Huynh-Pham +3
We propose an empirically stable and asymptotically efficient covariate-balancing approach to the problem of estimating survival causal effects in data with conditionally-independe…