◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Junxian He

11 papers hereh-index 10830 citations14 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • last author11

Across the 11 of 11 papers where every author was matched, so the position is known.

fields
  • cs.CL5
  • cs.AI3
  • cs.LG2
  • cs.CV1
same name
  • Junxian He — 11 papers, h 21
  • Junxian He — 7 papers
  • Junxian He — 7 papers, h 6
  • Junxian He — 5 papers, h 5
  • Junxian He — 4 papers
  • Junxian He — 3 papers, h 10

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
most citedWebExplorer: Explore and Evolve for Training Long-Horizon Web Agents

1 citations · 1 across the 5 of their papers we have counts for

collaborators
Showing cs.LGShow all

3 papers · 1 filter

cs.LG2025

Mirage or Method? How Model-Task Alignment Induces Divergent RL Conclusions

Haoze Wu, Cheng Wang, Wenshuo Zhao +1

Recent advances in applying reinforcement learning (RL) to large language models (LLMs) have led to substantial progress. In particular, a series of remarkable yet often counterint…

cs.LG2025

From Accuracy to Robustness: A Study of Rule- and Model-based Verifiers in Mathematical Reasoning

Yuzhen Huang, Weihao Zeng, Xingshan Zeng +2

Trustworthy verifiers are essential for the success of reinforcement learning with verifiable reward (RLVR), which is the core methodology behind various large reasoning models suc…

cs.LG2025

SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Weihao Zeng, Yuzhen Huang, Qian Liu +4

DeepSeek-R1 has shown that long chain-of-thought (CoT) reasoning can naturally emerge through a simple reinforcement learning (RL) framework with rule-based rewards, where the trai…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.