◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Anuj Mahajan

2 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author1

Across the 2 of 2 papers where every author was matched, so the position is known.

fields
  • cs.LG2
ORCID 0000-0003-2229-5287
same name
  • Anuj Mahajan — 10 papers, h 12

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedTrust-Region-Free Policy Optimization for Stochastic Policies

1 citations · 1 across the 2 of their papers we have counts for

collaborators

2 papers

cs.LG2023

Generalization Across Observation Shifts in Reinforcement Learning

Anuj Mahajan, Amy Zhang

Learning policies which are robust to changes in the environment are critical for real world deployment of Reinforcement Learning agents. They are also necessary for achieving good…

cs.LG2023★ 1 cited

Trust-Region-Free Policy Optimization for Stochastic Policies

Mingfei Sun, Benjamin Ellis, Anuj Mahajan +3

Trust Region Policy Optimization (TRPO) is an iterative method that simultaneously maximizes a surrogate objective and enforces a trust region constraint over consecutive policies…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.