◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Wei Yang

17 papers hereh-index 181.4k citations30 works total

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • middle author11
  • last author3

Across the 14 of 17 papers where every author was matched, so the position is known.

fields
  • cs.AI6
  • cs.LG6
  • cs.CL1
  • cs.CV1
  • cs.GT1
  • cs.MA1
same name
  • Wei Yang — 29 papers, h 12
  • Wei Yang — 21 papers, h 24
  • Wei Yang — 14 papers, h 4
  • Wei Yang — 13 papers, h 30
  • Wei Yang — 10 papers, h 16
  • Wei Yang — 10 papers, h 6

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

activity
20182023
most citedSupervised Learning Achieves Human-Level Performance in MOBA Games: A Case Study of Honor of Kings

55 citations · 121 across the 11 of their papers we have counts for

collaborators
Showing 2023Show all

4 papers · 1 filter

cs.AI2023

RLTF: Reinforcement Learning from Unit Test Feedback

Jiate Liu, Yiqin Zhu, Kaiwen Xiao +4

The goal of program synthesis, or code generation, is to generate executable code based on given descriptions. Recently, there has been an increasing number of studies employing re…

cs.GT2023

Policy Space Diversity for Non-Transitive Games

Jian Yao, Weiming Liu, Haobo Fu +4

Policy-Space Response Oracles (PSRO) is an influential algorithm framework for approximating a Nash Equilibrium (NE) in multi-agent non-transitive games. Many previous studies have…

cs.LG2023★ 2 cited

Future-conditioned Unsupervised Pretraining for Decision Transformer

Zhihui Xie, Zichuan Lin, Deheng Ye +3

Recent research in offline reinforcement learning (RL) has demonstrated that return-conditioned supervised learning is a powerful paradigm for decision-making problems. While promi…

cs.CL2023

Dynamic Transformers Provide a False Sense of Efficiency

Yiming Chen, Simin Chen, Zexin Li +4

Despite much success in natural language processing (NLP), pre-trained language models typically lead to a high computational cost during inference. Multi-exit is a mainstream appr…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.