5 papers
Active World-Model with 4D-informed Retrieval for Exploration and Awareness
Elaheh Vaezpour, Amirhosein Javadi, Tara Javidi
Physical awareness, especially in a large and dynamic environment, is shaped by sensing decisions that determine observability across space, time, and scale, while observations imp…
Convex and Non-convex Federated Learning with Stale Stochastic Gradients: Diminishing Step Size is All You Need
Xinran Zheng, Tara Javidi, Behrouz Touri
We propose a general framework for distributed stochastic optimization under delayed gradient models. In this setting, local agents leverage their own data and computation to a…
-Explorer: A Unified Framework for Active Model Estimation in MDPs
Xihe Gu, Urbashi Mitra, Tara Javidi
In tabular Markov decision processes (MDPs) with perfect state observability, each trajectory provides active samples from the transition distributions conditioned on state-action…
From Relative Entropy to Minimax: A Unified Framework for Coverage in MDPs
Xihe Gu, Urbashi Mitra, Tara Javidi
Targeted and deliberate exploration of state--action pairs is essential in reward-free Markov Decision Problems (MDPs). More precisely, different state-action pairs exhibit differe…
Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning
Talha Bozkus, Tara Javidi, Urbashi Mitra
Q-learning is widely employed for optimizing various large-dimensional networks with unknown system dynamics. Recent advancements include multi-environment mixed Q-learning (MEMQ)…