1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.AI2025★ 1 cited
Reinforcement Learning with Rubric Anchors
Zenan Huang, Yihong Zhuang, Guoshan Lu +18
Reinforcement Learning from Verifiable Rewards (RLVR) has emerged as a powerful paradigm for enhancing Large Language Models (LLMs), exemplified by the success of OpenAI's o-series…
cs.CV2023
CAME: Contrastive Automated Model Evaluation
Ru Peng, Qiuyang Duan, Haobo Wang +5
The Automated Model Evaluation (AutoEval) framework entertains the possibility of evaluating a trained machine learning model without resorting to a labeled testing set. Despite th…
cs.CL2023
Better Sign Language Translation with Monolingual Data
Ru Peng, Yawen Zeng, Junbo Zhao
Sign language translation (SLT) systems, which are often decomposed into video-to-gloss (V2G) recognition and gloss-to-text (G2T) translation through the pivot gloss, heavily relie…