collaborators

5 papers

cs.CV2026

Active World-Model with 4D-informed Retrieval for Exploration and Awareness

Elaheh Vaezpour, Amirhosein Javadi, Tara Javidi

Physical awareness, especially in a large and dynamic environment, is shaped by sensing decisions that determine observability across space, time, and scale, while observations imp…

math.OC2026

Convex and Non-convex Federated Learning with Stale Stochastic Gradients: Diminishing Step Size is All You Need

Xinran Zheng, Tara Javidi, Behrouz Touri

We propose a general framework for distributed stochastic optimization under delayed gradient models. In this setting, local agents leverage their own data and computation to a…

cs.LG2026

-Explorer: A Unified Framework for Active Model Estimation in MDPs

Xihe Gu, Urbashi Mitra, Tara Javidi

In tabular Markov decision processes (MDPs) with perfect state observability, each trajectory provides active samples from the transition distributions conditioned on state-action…

cs.LG2026

From Relative Entropy to Minimax: A Unified Framework for Coverage in MDPs

Xihe Gu, Urbashi Mitra, Tara Javidi

Targeted and deliberate exploration of state--action pairs is essential in reward-free Markov Decision Problems (MDPs). More precisely, different state-action pairs exhibit differe…

cs.LG2024

Coverage Analysis for Digital Cousin Selection -- Improving Multi-Environment Q-Learning

Talha Bozkus, Tara Javidi, Urbashi Mitra

Q-learning is widely employed for optimizing various large-dimensional networks with unknown system dynamics. Recent advancements include multi-environment mixed Q-learning (MEMQ)…