2 papers
cs.CL2025
Soft Token Attacks Cannot Reliably Audit Unlearning in Large Language Models
Haokun Chen, Sebastian Szyller, Weilin Xu +1
Large language models (LLMs) are trained using massive datasets, which often contain undesirable content such as harmful texts, personal information, and copyrighted material. To a…
cs.CV2024
Imperceptible Adversarial Examples in the Physical World
Weilin Xu, Sebastian Szyller, Cory Cornelius +5
Adversarial examples in the digital domain against deep learning-based computer vision models allow for perturbations that are imperceptible to human eyes. However, producing simil…