3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CL2024
Towards Generating Informative Textual Description for Neurons in Language Models
Shrayani Mondal, Rishabh Garodia, Arbaaz Qureshi +2
Recent developments in transformer-based language models have allowed them to capture a wide variety of world knowledge that can be adapted to downstream tasks with limited resourc…
cs.LG2023★ 3 cited
URET: Universal Robustness Evaluation Toolkit (for Evasion)
Kevin Eykholt, Taesung Lee, Douglas Schales +3
Machine learning models are known to be vulnerable to adversarial evasion attacks as illustrated by image classification models. Thoroughly understanding such attacks is critical i…