2 papers
cs.CL2026
APEX: Automated Prompt Engineering eXpert with Dynamic Data Selection
Fei Wang, Si Si, Cho-Jui Hsieh +1
Large Language Models are highly sensitive to prompt formulation, necessitating automatic prompt optimization to unlock their full potential. While evolutionary algorithms have eme…
cs.CL2025
BadJudge: Backdoor Vulnerabilities of LLM-as-a-Judge
Terry Tong, Fei Wang, Zhe Zhao +1
This paper proposes a novel backdoor threat attacking the LLM-as-a-Judge evaluation regime, where the adversary controls both the candidate and evaluator model. The backdoored eval…