◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Chao Yang

12 papers hereh-index 5130 citations17 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author7
  • last author4

Across the 12 of 12 papers where every author was matched, so the position is known.

fields
  • cs.CL4
  • cs.LG4
  • cs.CV2
  • cs.AI1
  • cs.CY1
same name
  • Chao Yang — 26 papers, h 49
  • Chao Yang — 16 papers, h 21
  • Chao Yang — 11 papers, h 4
  • Chao Yang — 10 papers, h 18
  • Chao Yang — 10 papers, h 8
  • Chao Yang — 9 papers, h 3

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators
Showing cs.LGShow all

4 papers · 1 filter

cs.LG2026

Subliminal Learning as Trait-Direction Drift: A Mechanism and Targeted Control under SFT Distillation

Zhixuan Liu, Zhichen Dong, Yuyu Fan +2

Beyond intended capabilities, model distillation can transfer hidden traits from a teacher. A teacher biased by a system prompt can generate semantically clean training data, such…

cs.LG2026

Scaling Model-Generated Distillation Data Can Make Latent Teacher Traits More Recoverable

Zhichen Dong, Zhixuan Liu, Yuyu Fan +3

Scaling model-generated data is usually viewed as improving distillation: more examples should increase coverage, reduce noise, and produce stronger students. We show a second effe…

cs.LG2026

Native Reasoning Models: Training Language Models to Reason on Unverifiable Data

Yuanfu Wang, Zhixuan Liu, Xiangtian Li +2

The prevailing paradigm for training large reasoning models--combining Supervised Fine-Tuning (SFT) with Reinforcement Learning with Verifiable Rewards (RLVR)--is fundamentally con…

cs.LG2025

Adversarial Preference Learning for Robust LLM Alignment

Yuanfu Wang, Pengyu Wang, Chenyang Xi +13

Modern language models often rely on Reinforcement Learning from Human Feedback (RLHF) to encourage safe behaviors. However, they remain vulnerable to adversarial attacks due to th…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.