2 papers
cs.LG2026
-Explorer: A Unified Framework for Active Model Estimation in MDPs
Xihe Gu, Urbashi Mitra, Tara Javidi
In tabular Markov decision processes (MDPs) with perfect state observability, each trajectory provides active samples from the transition distributions conditioned on state-action…
cs.LG2025
ModShift: Model Privacy via Designed Shifts
Nomaan A. Kherani, Urbashi Mitra
In this paper, shifts are introduced to preserve model privacy against an eavesdropper in federated learning. Model learning is treated as a parameter estimation problem. This pers…