2 papers
cs.CL2023
LimeAttack: Local Explainable Method for Textual Hard-Label Adversarial Attack
Hai Zhu, Zhaoqing Yang, Weiwei Shang +1
Natural language processing models are vulnerable to adversarial examples. Previous textual adversarial attacks adopt gradients or confidence scores to calculate word importance ra…
cs.CL2023
BeamAttack: Generating High-quality Textual Adversarial Examples through Beam Search and Mixed Semantic Spaces
Hai Zhu, Qingyang Zhao, Yuren Wu
Natural language processing models based on neural networks are vulnerable to adversarial examples. These adversarial examples are imperceptible to human readers but can mislead mo…