3 papers
cs.CL2026
APEX: Automated Prompt Engineering eXpert with Dynamic Data Selection
Fei Wang, Si Si, Cho-Jui Hsieh +1
Large Language Models are highly sensitive to prompt formulation, necessitating automatic prompt optimization to unlock their full potential. While evolutionary algorithms have eme…
cs.CL2026
GroupDPO: Memory efficient Group-wise Direct Preference Optimization
Jixuan Leng, Si Si, Hsiang-Fu Yu +2
Preference optimization is widely used to align Large Language Models (LLMs) with preference feedback. However, most existing methods train on a single positive-negative pair per p…
cs.LG2025
LoRA Done RITE: Robust Invariant Transformation Equilibration for LoRA Optimization
Jui-Nan Yen, Si Si, Zhao Meng +5
Low-rank adaption (LoRA) is a widely used parameter-efficient finetuning method for LLM that reduces memory requirements. However, current LoRA optimizers lack transformation invar…