◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Huaijie Wang

5 papers hereh-index 3311 citations7 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author2
  • middle author3

Across the 5 of 5 papers where every author was matched, so the position is known.

fields
  • cs.LG2
  • cs.AI1
  • cs.DC1
  • cs.RO1
same name
  • Huaijie Wang — 7 papers, h 3
  • Huaijie Wang — 4 papers, h 1

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

collaborators

5 papers

cs.DC2026

Next-Generation Agentic Reinforcement Learning Systems Enable Self-Evolving Agents

Ran Yan, Wei Fu, Jiale Li +21

LLM agents are rapidly being deployed in production, including coding assistants, customer-support chatbots, and scientific research assistants, yet they remain fundamentally stati…

cs.LG2026

Building Multi-Task Agentic LLMs via Two-Phase Distillation

Huaijie Wang, Shusheng Xu, Yi Wu +1

A key step toward artificial general intelligence is to train models that can perform multiple tasks. In this paper, we study how to build such models by first training separate RL…

cs.AI2026

Verifiable Process Rewards for Agentic Reasoning

Huining Yuan, Zelai Xu, Huaijie Wang +6

Reinforcement learning from verifiable rewards (RLVR) has improved the reasoning abilities of large language models (LLMs), but most existing approaches rely on sparse outcome-leve…

cs.LG2024

Offline Reinforcement Learning for LLM Multi-Step Reasoning

Huaijie Wang, Shibo Hao, Hanze Dong +4

Improving the multi-step reasoning ability of large language models (LLMs) with offline reinforcement learning (RL) is essential for quickly adapting them to complex tasks. While D…

cs.RO2024

LAGOON: Language-Guided Motion Control

Shusheng Xu, Huaijie Wang, Jiaxuan Gao +3

We aim to control a robot to physically behave in the real world following any high-level language command like "cartwheel" or "kick". Although human motion datasets exist, this ta…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.