4 papers
MIRAGE: Protecting against Malicious Image Editing via False Moderation
Anshul Nasery, Ramnath Kumar, Cho-Jui Hsieh +1
The proliferation of AI-powered image editing systems raises serious concerns because it allows personal images to be arbitrarily manipulated at scale, with minimal effort, and a l…
DualEval: Joint Model-Item Calibration for Unified LLM Evaluation
Aaron J. Li, Hao Huang, Youngmin Park +6
Current LLM evaluation relies on two complementary but often disconnected signals: static benchmarks with objective correctness labels and arena-style preference data that better r…
Cycle-Consistent Search: Question Reconstructability as a Proxy Reward for Search Agent Training
Sohyun An, Shuibenyang Yuan, Hayeon Lee +2
Reinforcement Learning (RL) has shown strong potential for optimizing search agents in complex information retrieval tasks. However, existing approaches predominantly rely on gold…
Provably Robust Training of Quantum Circuit Classifiers Against Parameter Noise
Lucas Tecot, Di Luo, Cho-Jui Hsieh
Advancements in quantum computing have spurred significant interest in harnessing its potential for speedups over classical systems. However, noise remains a major obstacle to achi…