◍wovepaper
SearchResearchersInstitutions
Sign in
researcher

Youpeng Zhao

4 papers here

Matching runs newest-first, so older work may not be attached to this profile yet.

author position
  • first author1
  • middle author2

Across the 3 of 4 papers where every author was matched, so the position is known.

fields
  • cs.AI2
  • cs.LG1
  • cs.MA1
ORCID 0000-0002-4610-3545
same name
  • Youpeng Zhao — 4 papers
  • Youpeng Zhao — 1 paper

Either other researchers who publish under this name, or the same person where the external sources have not merged their records.

identity via Semantic Scholar / OpenAlex

most citedALISA: Accelerating Large Language Model Inference via Sparsity-Aware KV Caching

1 citations · 1 across the 4 of their papers we have counts for

collaborators

4 papers

cs.AI2024

RL-LLM-DT: An Automatic Decision Tree Generation Method Based on RL Evaluation and LLM Enhancement

Junjie Lin, Jian Zhao, Lin Liu +6

Traditionally, AI development for two-player zero-sum games has relied on two primary techniques: decision trees and reinforcement learning (RL). A common approach involves using a…

cs.LG2024

CuDA2: An approach for Incorporating Traitor Agents into Cooperative Multi-Agent Systems

Zhen Chen, Yong Liao, Youpeng Zhao +2

Cooperative Multi-Agent Reinforcement Learning (CMARL) strategies are well known to be vulnerable to adversarial perturbations. Previous works on adversarial attacks have primarily…

cs.MA2024

Mini Honor of Kings: A Lightweight Environment for Multi-Agent Reinforcement Learning

Lin Liu, Jian Zhao, Cheng Hu +9

Games are widely used as research environments for multi-agent reinforcement learning (MARL), but they pose three significant challenges: limited customization, high computational…

cs.AI2024★ 1 cited

ALISA: Accelerating Large Language Model Inference via Sparsity-Aware KV Caching

Youpeng Zhao, Di Wu, Jun Wang

The Transformer architecture has significantly advanced natural language processing (NLP) and has been foundational in developing large language models (LLMs) such as LLaMA and OPT…

◍wovepaper

Papers, researchers and institutions, woven together.

Explore
  • Search
  • Researchers
  • Institutions
Account
  • Library
  • Chat
Data
  • arXiv.org
  • Semantic Scholar
  • OpenAlex
  • Latest RSS
AboutContactPrivacyDevelopersllms.txtopenapi.json
Not affiliated with arXiv. Researcher data from Semantic Scholar (ODC-BY) and OpenAlex.