1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.AI2026
MCPO: Mastery-Consolidated Policy Optimization for Large Reasoning Models
Zhaokang Liao, Yingguo Gao, Yi Yang +2
Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a promising approach to improve the reasoning abilities of Large Language Models (LLMs). Among RLVR algorithms,…
cs.LG2022
Facial Affect Analysis: Learning from Synthetic Data & Multi-Task Learning Challenges
Siyang Li, Yifan Xu, Huanyu Wu +4
Facial affect analysis remains a challenging task with its setting transitioned from lab-controlled to in-the-wild situations. In this paper, we present novel frameworks to handle…
cs.CV2022★ 1 cited
Cross-Architecture Knowledge Distillation
Yufan Liu, Jiajiong Cao, Bing Li +3
Transformer attracts much attention because of its ability to learn global relations and superior performance. In order to achieve higher performance, it is natural to distill comp…