2 citations · 2 across the 2 of their papers we have counts for
3 papers
cs.CL2025
AURORA:Automated Training Framework of Universal Process Reward Models via Ensemble Prompting and Reverse Verification
Xiaoyu Tan, Tianchu Yao, Chao Qu +8
The reasoning capabilities of advanced large language models (LLMs) like o1 have revolutionized artificial intelligence applications. Nevertheless, evaluating and optimizing comple…
cs.CL2025
Refine Knowledge of Large Language Models via Adaptive Contrastive Learning
Yinghui Li, Haojing Huang, Jiayi Kuang +7
How to alleviate the hallucinations of Large Language Models (LLMs) has always been the fundamental goal pursued by the LLMs research community. Looking through numerous hallucinat…
cs.LG2024★ 2 cited
Adaptive Learning on User Segmentation: Universal to Specific Representation via Bipartite Neural Interaction
Xiaoyu Tan, Yongxin Deng, Chao Qu +4
Recently, models for user representation learning have been widely applied in click-through-rate (CTR) and conversion-rate (CVR) prediction. Usually, the model learns a universal u…