◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Rahul Singh

4 papers hereh-index 18 citations11 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • last author2

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.LG3
  • cs.ET1
same name
  • Rahul Singh — 11 papers, h 6
  • Rahul Singh — 4 papers, h 4
  • Rahul Singh — 3 papers, h 6
  • Rahul Singh — 3 papers, h 1
  • Rahul Singh — 3 papers, h 1
  • Rahul Singh — 2 papers, h 1

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

4 papers

cs.LG2026

Regret and Sample Complexity of Online Q-Learning via Concentration of Stochastic Approximation with Time-Inhomogeneous Markov Chains

Rahul Singh, Siddharth Chandak, Eric Moulines +2

We present the first regret bound for classical online Q-learning in infinite-horizon discounted Markov decision processes (MDPs), without relying on optimism or bonus terms. We fi…

cs.ET2026

Robustness Verification of Binary Neural Networks: An Ising and Quantum-Inspired Framework

Rahul Singh, Seyran Saeedi, Zheng Zhang

Binary neural networks (BNNs) are increasingly deployed in edge computing applications due to their low hardware complexity and high energy efficiency. However, verifying the robus…

cs.LG2025

Policy Zooming: Adaptive Discretization-based Infinite-Horizon Average-Reward Reinforcement Learning

Avik Kar, Rahul Singh

We study the infinite-horizon average-reward reinforcement learning (RL) for continuous space Lipschitz MDPs in which an agent can play policies from a given set I^¦. The proposed…

cs.LG2025

Provably Adaptive Average Reward Reinforcement Learning for Metric Spaces

Avik Kar, Rahul Singh

We study infinite-horizon average-reward reinforcement learning (RL) for Lipschitz MDPs, a broad class that subsumes several important classes such as linear and RKHS MDPs, functio…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.