◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Yaodong Yang

4 papers hereh-index 124 citations5 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author1
  • last author2

Across the 3 of 4 papers where every author was matched, so the position is known.

fields
  • cs.CL1
  • cs.LG1
  • cs.MA1
  • cs.RO1
same name
  • Yaodong Yang — 39 papers, h 18
  • Yaodong Yang — 15 papers, h 8
  • Yaodong Yang — 10 papers, h 3
  • Yaodong Yang — 6 papers, h 2
  • Yaodong Yang — 6 papers, h 3
  • Yaodong Yang — 5 papers, h 6

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20242026
collaborators

4 papers

cs.LG2026

Finding Kissing Numbers with Game-theoretic Reinforcement Learning

Chengdong Ma, Théo Tao Zhaowei, Pengyu Li +7

Since Isaac Newton first studied the Kissing Number Problem in 1694, determining the maximal number of non-overlapping spheres around a central sphere has remained a defining chall…

cs.RO2026

Accelerating Robotic Reinforcement Learning with Agent Guidance

Haojun Chen, Zili Zou, Chengdong Ma +4

Reinforcement Learning (RL) offers a powerful paradigm for autonomous robots to master generalist manipulation skills through trial-and-error. However, its real-world application i…

cs.CL2025

Magnetic Preference Optimization: Achieving Last-iterate Convergence for Language Model Alignment

Mingzhi Wang, Chengdong Ma, Qizhi Chen +7

Self-play methods have demonstrated remarkable success in enhancing model capabilities across various domains. In the context of Reinforcement Learning from Human Feedback (RLHF),…

cs.MA2024

Towards Efficient Collaboration via Graph Modeling in Reinforcement Learning

Wenzhe Fan, Zishun Yu, Chengdong Ma +3

In multi-agent reinforcement learning, a commonly considered paradigm is centralized training with decentralized execution. However, in this framework, decentralized execution rest…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.