2 papers
cs.LG2026
AdaPrivate-TS: Private Thompson Sampling for Contextual Bandits with Privacy Amplification
Mohammadreza Riyazat, Eranga Ukwatta
We present AdaPrivate-TS, a differentially private contextual bandit algorithm that combines Thompson Sampling with batched zCDP composition. Our key insight is that differential p…
cs.CL2026
MARD: Mirror-Augmented Reasoning Distillation for Mechanism-Level Drug-Drug Interaction Prediction
Mohammadreza Riyazat, Vian Lelo, Rameen Jafri +2
Mechanism-level drug-drug interaction (DDI) prediction requires identifying which enzyme or pharmacodynamic axis is implicated, in which direction, and with which evidence -- not m…