3 papers
cs.CL2026
SWAN: Semantic Watermarking with Abstract Meaning Representation
Ziping Ye, Gourab Dey, Christos Christodoulopoulos +7
We introduce SWAN (Semantic Watermarking with Abstract Meaning Representation), a novel framework that embeds watermark signatures into the semantic structure of a sentence using A…
cs.CR2026
Attacks Meet Interpretability (AmI) Evaluation and Findings
Qian Ma, Ziping Ye, Shagufta Mehnaz
To investigate the effectiveness of the model explanation in detecting adversarial examples, we reproduce the results of two papers, Attacks Meet Interpretability: Attribute-steere…
cs.CR2025
Enhancing Adversarial Example Detection Through Model Explanation
Qian Ma, Ziping Ye
Adversarial examples are a major problem for machine learning models, leading to a continuous search for effective defenses. One promising direction is to leverage model explanatio…