◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Sheng Jia

5 papers hereh-index 3184 citations6 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author3
  • middle author1

Across the 4 of 5 papers where every author was matched, so the position is known.

fields
  • cs.CL2
  • cs.LG2
  • cs.AI1

identity via Semantic Scholar / OpenAlex

activity
20192026
most citedDOM-Q-NET: Grounded RL on Structured Language

9 citations · 9 across the 4 of their papers we have counts for

collaborators
Showing cs.LGShow all

2 papers · 1 filter

cs.LG2026

Group Adaptive Clipping Policy Optimization

Sheng Jia, Xiao Wang, Shiva Prasad Kasiviswanathan +1

Group relative policy optimization for reinforcement learning with verifiable rewards (RLVR) typically uses a fixed importance-sampling (IS) ratio clipping boundary across all roll…

cs.LG2019★ 9 cited

DOM-Q-NET: Grounded RL on Structured Language

Sheng Jia, Jamie Kiros, Jimmy Ba

Building agents to interact with the web would allow for significant improvements in knowledge understanding and representation learning. However, web navigation tasks are difficul…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.