◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Jiajun Zhang

4 papers hereh-index 3130 citations5 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • last author4

Across the 4 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CL3
  • cs.AI1
same name
  • Jiajun Zhang — 30 papers, h 39
  • Jiajun Zhang — 17 papers, h 14
  • Jiajun Zhang — 16 papers
  • Jiajun Zhang — 10 papers, h 13
  • Jiajun Zhang — 7 papers, h 4
  • Jiajun Zhang — 7 papers, h 5

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.CLShow all

4 papers · 1 filter

cs.CL2026

EDGE: Experience-Distillation for Guided Exploration in Agentic Reinforcement Learning

Can Xie, Yuyi Zhou, Wen Yang +5

Reinforcement learning with outcome-based objectives such as GRPO enables LLM-based agents to solve complex, long-horizon tasks, yet the reusable exploration patterns embedded in i…

cs.CL2025

An Efficient and Precise Training Data Construction Framework for Process-supervised Reward Model in Mathematical Reasoning

Wei Sun, Qianlong Du, Fuwei Cui +1

Enhancing the mathematical reasoning capabilities of Large Language Models (LLMs) is of great scientific and practical significance. Researchers typically employ process-supervised…

cs.CL2024

ChineseWebText 2.0: Large-Scale High-quality Chinese Web Text with Multi-dimensional and fine-grained information

Wanyue Zhang, Ziyong Li, Wen Yang +5

During the development of large language models (LLMs), pre-training data play a critical role in shaping LLMs' capabilities. In recent years several large-scale and high-quality p…

cs.CL2024

A Survey on Data Selection for LLM Instruction Tuning

Bolin Zhang, Jiahao Wang, Qianlong Du +3

Instruction tuning is a vital step of training large language models (LLMs), so how to enhance the effect of instruction tuning has received increased attention. Existing works ind…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.