4 papers
Trojans in Artificial Intelligence (TrojAI) Final Report
Kristopher W. Reese, Taylor Kulp-McDowall, Michael Majurski +68
The Intelligence Advanced Research Projects Activity (IARPA) launched the TrojAI program to confront an emerging vulnerability in modern artificial intelligence: the threat of AI T…
-Explorer: A Unified Framework for Active Model Estimation in MDPs
Xihe Gu, Urbashi Mitra, Tara Javidi
In tabular Markov decision processes (MDPs) with perfect state observability, each trajectory provides active samples from the transition distributions conditioned on state-action…
From Relative Entropy to Minimax: A Unified Framework for Coverage in MDPs
Xihe Gu, Urbashi Mitra, Tara Javidi
Targeted and deliberate exploration of state--action pairs is essential in reward-free Markov Decision Problems (MDPs). More precisely, different state-action pairs exhibit differe…
Trojan Cleansing with Neural Collapse
Xihe Gu, Greg Fields, Yaman Jandali +2
Trojan attacks are sophisticated training-time attacks on neural networks that embed backdoor triggers which force the network to produce a specific output on any input which inclu…