activity
20222025
most citedLightweight Vision Transformer with Cross Feature Attention

8 citations · 13 across the 11 of their papers we have counts for

collaborators
Showing 2022Show all

6 papers · 1 filter

cs.AI2022

DanZero: Mastering GuanDan Game with Reinforcement Learning

Yudong Lu, Jian Zhao, Youpeng Zhao +2

Card game AI has always been a hot topic in the research of artificial intelligence. In recent years, complex card games such as Mahjong, DouDizhu and Texas Hold'em have been solve…

cs.CV2022★ 8 cited

Lightweight Vision Transformer with Cross Feature Attention

Youpeng Zhao, Huadong Tang, Yingying Jiang +2

Recent advances in vision transformers (ViTs) have achieved great performance in visual recognition tasks. Convolutional neural networks (CNNs) exploit spatial inductive bias to le…

cs.CV2022★ 3 cited

Semi-Supervised Segmentation of Mitochondria from Electron Microscopy Images Using Spatial Continuity

Yunpeng Xiao, Youpeng Zhao, Ge Yang

Morphology of mitochondria plays critical roles in mediating their physiological functions. Accurate segmentation of mitochondria from 3D electron microscopy (EM) images is essenti…

cs.AI2022★ 1 cited

DouZero+: Improving DouDizhu AI by Opponent Modeling and Coach-guided Learning

Youpeng Zhao, Jian Zhao, Xunhan Hu +2

Recent years have witnessed the great breakthrough of deep reinforcement learning (DRL) in various perfect and imperfect information games. Among these games, DouDizhu, a popular c…

cs.LG2022★ 1 cited

Coach-assisted Multi-Agent Reinforcement Learning Framework for Unexpected Crashed Agents

Jian Zhao, Youpeng Zhao, Weixun Wang +5

Multi-agent reinforcement learning is difficult to be applied in practice, which is partially due to the gap between the simulated and real-world scenarios. One reason for the gap…

cs.LG2022

MCMARL: Parameterizing Value Function via Mixture of Categorical Distributions for Multi-Agent Reinforcement Learning

Jian Zhao, Mingyu Yang, Youpeng Zhao +4

In cooperative multi-agent tasks, a team of agents jointly interact with an environment by taking actions, receiving a team reward and observing the next state. During the interact…