◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yang Zhang

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author2

Across the 2 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CV1
  • cs.LG1
  • cs.MA1
  • cs.RO1
same name
  • Yang Zhang — 28 papers, h 18
  • Yang Zhang — 24 papers, h 11
  • Yang Zhang — 22 papers
  • Yang Zhang — 17 papers, h 27
  • Yang Zhang — 17 papers
  • Yang Zhang — 16 papers, h 12

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedOnline Iterative Self-Alignment for Radiology Report Generation

1 citations · 1 across the 3 of their papers we have counts for

collaborators

4 papers

cs.RO2025

Align-Then-stEer: Adapting the Vision-Language Action Models through Unified Latent Guidance

Yang Zhang, Chenwei Wang, Ouyang Lu +7

Vision-Language-Action (VLA) models pre-trained on large, diverse datasets show remarkable potential for general-purpose robotic manipulation. However, a primary bottleneck remains…

cs.CV2025★ 1 cited

Online Iterative Self-Alignment for Radiology Report Generation

Ting Xiao, Lei Shi, Yang Zhang +3

Radiology Report Generation (RRG) is an important research topic for relieving radiologist' heavy workload. Existing RRG models mainly rely on supervised fine-tuning (SFT) based on…

cs.MA2025

Revisiting Multi-Agent World Modeling from a Diffusion-Inspired Perspective

Yang Zhang, Xinran Li, Jianing Ye +5

World models have recently attracted growing interest in Multi-Agent Reinforcement Learning (MARL) due to their ability to improve sample efficiency for policy learning. However, a…

cs.LG2025

Online Preference Alignment for Language Models via Count-based Exploration

Chenjia Bai, Yang Zhang, Shuang Qiu +3

Reinforcement Learning from Human Feedback (RLHF) has shown great potential in fine-tuning Large Language Models (LLMs) to align with human preferences. Existing methods perform pr…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.