2 papers
cs.LG2025
Reward Hacking Mitigation using Verifiable Composite Rewards
Mirza Farhan Bin Tarek, Rahmatollah Beheshti
Reinforcement Learning from Verifiable Rewards (RLVR) has recently shown that large language models (LLMs) can develop their own reasoning without direct supervision. However, appl…
cs.LG2024
Fairness-Optimized Synthetic EHR Generation for Arbitrary Downstream Predictive Tasks
Mirza Farhan Bin Tarek, Raphael Poulain, Rahmatollah Beheshti
Among various aspects of ensuring the responsible design of AI tools for healthcare applications, addressing fairness concerns has been a key focus area. Specifically, given the wi…