3 papers
cs.LG2026
Efficient Hyperparameter Optimization for LLM Reinforcement Learning
Minping Chen, Bowen Xiao, Du Liang +2
Reinforcement learning (RL) for large language models (LLMs) is highly sensitive to hyperparameter configurations, making hyperparameter optimization (HPO) essential yet computatio…
cs.AI2026
Enhancing Online Recruitment with Category-Aware MoE and LLM-based Data Augmentation
Minping Chen, Bing Xu, Zulong Chen +4
Person-Job Fit (PJF) is a critical component for online recruitment. Existing approaches face several challenges, particularly in handling low-quality job descriptions and similar…
cs.IR2025
Enhancing Talent Search Ranking with Role-Aware Expert Mixtures and LLM-based Fine-Grained Job Descriptions
Jihang Li, Bing Xu, Zulong Chen +5
Talent search is a cornerstone of modern recruitment systems, yet existing approaches often struggle to capture nuanced job-specific preferences, model recruiter behavior at a fine…