◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

S. R. Eshwar

3 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • sole author1
  • first author2

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • cs.AI1

identity via Semantic Scholar / OpenAlex

collaborators

3 papers

cs.AI2025

Reliable Policy Iteration: Performance Robustness Across Architecture and Environment Perturbations

S. R. Eshwar, Aniruddha Mukherjee, Kintan Saha +4

In a recent work, we proposed Reliable Policy Iteration (RPI), that restores policy iteration's monotonicity-of-value-estimates property to the function approximation setting. Here…

cs.LG2025

Teaching Precommitted Agents: Model-Free Policy Evaluation and Control in Quasi-Hyperbolic Discounted MDPs

S. R. Eshwar

Time-inconsistent preferences, where agents favor smaller-sooner over larger-later rewards, are a key feature of human and animal decision-making. Quasi-Hyperbolic (QH) discounting…

cs.LG2024

Reinforcement Learning with Quasi-Hyperbolic Discounting

S. R. Eshwar, Mayank Motwani, Nibedita Roy +1

Reinforcement learning has traditionally been studied with exponential discounting or the average reward setup, mainly due to their mathematical tractability. However, such framewo…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.