Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
APEX: Automated Prompt Engineering eXpert with Dynamic Data Selection
Fei Wang, Si Si, Cho-Jui Hsieh +1
Large Language Models are highly sensitive to prompt formulation, necessitating automatic prompt optimization to unlock their full potential. While evolutionary algorithms have eme…
cs.CL2026
GroupDPO: Memory efficient Group-wise Direct Preference Optimization
Jixuan Leng, Si Si, Hsiang-Fu Yu +2
Preference optimization is widely used to align Large Language Models (LLMs) with preference feedback. However, most existing methods train on a single positive-negative pair per p…