2 papers
cs.LG2026
Labeled TrustSet Guided: Batch Active Learning with Reinforcement Learning
Guofeng Cui, Yang Liu, Pichao Wang +4
Batch active learning (BAL) is a crucial technique for reducing labeling costs and improving data efficiency in training large-scale deep learning models. Traditional BAL methods o…
cs.CL2025
CRPO: Confidence-Reward Driven Preference Optimization for Machine Translation
Guofeng Cui, Pichao Wang, Yang Liu +3
Large language models (LLMs) have shown great potential in natural language processing tasks, but their application to machine translation (MT) remains challenging due to pretraini…