18 citations · 18 across the 3 of their papers we have counts for
3 papers
cs.LG2023
Black-Box Targeted Reward Poisoning Attack Against Online Deep Reinforcement Learning
Yinglun Xu, Gagandeep Singh
We propose the first black-box targeted attack against online deep reinforcement learning through reward poisoning during training time. Our attack is applicable to general environ…
cs.LG2023★ 18 cited
Incremental Verification of Neural Networks
Shubham Ugare, Debangshu Banerjee, Sasa Misailovic +1
Complete verification of deep neural networks (DNNs) can exactly determine whether the DNN satisfies a desired trustworthy property (e.g., robustness, fairness) on an infinite set…
cs.LG2023
Interpreting Robustness Proofs of Deep Neural Networks
Debangshu Banerjee, Avaljot Singh, Gagandeep Singh
In recent years numerous methods have been developed to formally verify the robustness of deep neural networks (DNNs). Though the proposed techniques are effective in providing mat…