2 papers
cs.LG2025
Robustness of deep learning classification to adversarial input on GPUs: asynchronous parallel accumulation is a source of vulnerability
Sanjif Shanmugavelu, Mathieu Taillefumier, Christopher Culver +3
The ability of machine learning (ML) classification models to resist small, targeted input perturbations -- known as adversarial attacks -- is a key measure of their safety and rel…
cs.DC2024
Impacts of floating-point non-associativity on reproducibility for HPC and deep learning applications
Sanjif Shanmugavelu, Mathieu Taillefumier, Christopher Culver +3
Run to run variability in parallel programs caused by floating-point non-associativity has been known to significantly affect reproducibility in iterative algorithms, due to accumu…