◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Vijay Gupta

4 papers hereh-index 27 citations5 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2
  • last author2

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.LG4
same name
  • Vijay Gupta — 4 papers, h 2
  • Vijay Gupta — 4 papers, h 1
  • Vijay Gupta — 3 papers, h 2
  • Vijay Gupta — 2 papers, h 1
  • Vijay Gupta — 2 papers, h 1
  • Vijay Gupta — 2 papers, h 1

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

4 papers

cs.LG2026

Learning to Control Coupled-Dynamics Environments with Joint Markov Decision Processes

Ege C. Kaya, Aliasghar Pourghani, Mahsa Ghasemi +2

Coupled-dynamics environments expose the one-step outcomes that would follow from several possible counterfactual actions under a common realization of exogenous randomness. The or…

cs.LG2026

Quotient-Categorical Representations for Bellman-Compatible Average-Reward Distributional Reinforcement Learning

Ege C. Kaya, Aliasghar Pourghani, Vijay Gupta +1

Average-reward reinforcement learning requires estimating the gain and the bias, which is defined only up to an additive constant. This makes direct distributional analogues ill-po…

cs.LG2026

Parameter-free Optimal Rates for Nonlinear Semi-Norm Contractions with Applications to Q-Learning

Ankur Naskar, Gugan Thoppe, Vijay Gupta

Algorithms for solving \textit{nonlinear} fixed-point equations -- such as average-reward \textit{Q-learning} and \textit{TD-learning} -- often involve semi-norm contractions. Ac…

cs.LG2025

Parameter-Free Federated TD Learning with Markov Noise in Heterogeneous Environments

Ankur Naskar, Gugan Thoppe, Utsav Negi +1

Federated learning (FL) can dramatically speed up reinforcement learning by distributing exploration and training across multiple agents. It can guarantee an optimal convergence ra…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.