2 papers
cs.CV2026
ALLUDE: A Unified Evaluation System for Configurable Attacks in Differentiable Environments
Mansi Phute, Alexander Greenhalgh, Matthew Hull +8
Adversarial attacks against vision models like object detectors are often evaluated under limited conditions, leaving their performance under-characterized. Bridging simulation and…
cs.HC2026
UNIPO: Unified Interactive Visual Explanation for RL Fine-Tuning Policy Optimization
Aeree Cho, Alexander D. Greenhalgh, Jonathan Bodea +2
Reinforcement learning has emerged as a dominant technique for fine-tuning the behavior of large language models, with policy optimization (PO) algorithms such as GRPO, DAPO, and D…