◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Srivatsan Srinivasan

3 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author3

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • cs.CL1
ORCID 0000-0002-6672-4779

identity via Semantic Scholar / OpenAlex

most citedReinforced Self-Training (ReST) for Language Modeling

18 citations · 23 across the 3 of their papers we have counts for

collaborators

3 papers

cs.CL2023★ 18 cited

Reinforced Self-Training (ReST) for Language Modeling

Caglar Gulcehre, Tom Le Paine, Srivatsan Srinivasan +11

Reinforcement learning from human feedback (RLHF) can improve the quality of large language model's (LLM) outputs by aligning them with human preferences. We propose a simple algor…

cs.LG2023★ 5 cited

AlphaStar Unplugged: Large-Scale Offline Reinforcement Learning

Michaël Mathieu, Sherjil Ozair, Srivatsan Srinivasan +21

StarCraft II is one of the most challenging simulated reinforcement learning environments; it is partially observable, stochastic, multi-agent, and mastering StarCraft II requires…

cs.LG2022

An Empirical Study of Implicit Regularization in Deep Offline RL

Caglar Gulcehre, Srivatsan Srinivasan, Jakub Sygnowski +5

Deep neural networks are the most commonly used function approximators in offline reinforcement learning. Prior works have shown that neural nets trained with TD-learning and gradi…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.