◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Raghav Jain

4 papers hereh-index 316 citations10 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author4

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.LG3
  • cs.MA1

identity via Semantic Scholar / OpenAlex

collaborators

4 papers

cs.LG2026

AI Assistants Overassist

Verona Teo, Raghav Jain, Tobias Gerstenberg +1

Large language models (LLMs) are increasingly used as tutors and thought partners, helping users reason through problems. While guidance from AI assistants can scaffold thinking an…

cs.LG2026

Generate, Filter, Control, Replay: A Comprehensive Survey of Rollout Strategies for LLM Reinforcement Learning

Rohan Surana, Gagan Mundada, Xunyi Jiang +19

Reinforcement learning (RL) has become a central post-training tool for improving the reasoning abilities of large language models (LLMs). In these systems, the rollout, the trajec…

cs.MA2026

The Subtle Art of Defection: Understanding Uncooperative Behaviors in LLM based Multi-Agent Systems

Devang Kulshreshtha, Wanyu Du, Raghav Jain +4

This paper introduces a novel framework for simulating and analyzing how uncooperative behaviors can destabilize or collapse LLM-based multi-agent systems. Our framework includes t…

cs.LG2025

Using Shapley interactions to understand how models use structure

Divyansh Singhvi, Diganta Misra, Andrej Erkelens +3

Language is an intricately structured system, and a key goal of NLP interpretability is to provide methodological insights for understanding how language models represent this stru…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.