3 papers
cs.LG2026
Smooth Multi-Policy Causal Effect Estimation in Longitudinal Settings
Wenxin Chen, Weishen Pan, Kyra Gan +1
Comparative evaluation of multiple dynamic treatment policies is essential for healthcare and policy decisions, yet conventional longitudinal causal inference methods estimate each…
cs.LG2026
ODRPO: Ordinal Decompositions of Discrete Rewards for Robust Policy Optimization
Nirmal Patel, Fei Wang, Inderjit S. Dhillon
The alignment of Large Language Models (LLMs) utilizes Reinforcement Learning from AI Feedback (RLAIF) for non-verifiable domains such as long-form question answering and open-ende…
cs.AI2023
Leveraging Generative AI for Clinical Evidence Summarization Needs to Ensure Trustworthiness
Gongbo Zhang, Qiao Jin, Denis Jered McInerney +11
Evidence-based medicine promises to improve the quality of healthcare by empowering medical decisions and practices with the best available evidence. The rapid growth of medical ev…