◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

S. R. Samineni

3 papers hereh-index 355 citations5 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • sole author1
  • first author2

Across the 3 of 3 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • cs.AI1

identity via Semantic Scholar / OpenAlex

collaborators

3 papers

cs.AI2025

Local Coherence or Global Validity? Investigating RLVR Traces in Math Domains

Soumya Rani Samineni, Durgesh Kalwar, Vardaan Gangal +2

Reinforcement Learning with Verifiable Rewards (RLVR)-based post-training of Large Language Models (LLMs) has been shown to improve accuracy on reasoning tasks and continues to att…

cs.LG2025

RL in Name Only? Analyzing the Structural Assumptions in RL post-training for LLMs

Soumya Rani Samineni, Durgesh Kalwar, Karthik Valmeekam +2

Reinforcement learning based post-training of large language models (LLMs) has recently gained attention, particularly following the release of DeepSeek R1, which applied GRPO for…

cs.LG2021

Policy Search using Dynamic Mirror Descent MPC for Model Free Off Policy RL

Soumya Rani Samineni

Recent works in Reinforcement Learning (RL) combine model-free (Mf)-RL algorithms with model-based (Mb)-RL approaches to get the best from both: asymptotic performance of Mf-RL and…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.