2 papers
cs.LG2026
Robust LLM Alignment via Distributionally Robust Direct Preference Optimization
Zaiyan Xu, Sushil Vemuri, Kishan Panaganti +3
A major challenge in aligning large language models (LLMs) with human preferences is the issue of distribution shift. LLM alignment algorithms rely on static preference datasets, a…
cs.CR2025
ReFuzz: Reusing Tests for Processor Fuzzing with Contextual Bandits
Chen Chen, Zaiyan Xu, Mohamadreza Rostami +4
Processor designs rely on iterative modifications and reuse well-established designs. However, this reuse of prior designs also leads to similar vulnerabilities across multiple pro…