Showing cs.LGShow all
3 papers · 1 filter
cs.LG2025
Universal and Transferable Adversarial Attack on Large Language Models Using Exponentiated Gradient Descent
Sajib Biswas, Mao Nishino, Samuel Jacob Chacko +1
As large language models (LLMs) are increasingly deployed in critical applications, ensuring their robustness and safety alignment remains a major challenge. Despite the overall su…
cs.LG2025
Adversarial Attack on Large Language Models using Exponentiated Gradient Descent
Sajib Biswas, Mao Nishino, Samuel Jacob Chacko +1
As Large Language Models (LLMs) are widely used, understanding them systematically is key to improving their safety and realizing their full potential. Although many models are ali…
cs.LG2024
Adversarial Attacks on Large Language Models Using Regularized Relaxation
Samuel Jacob Chacko, Sajib Biswas, Chashi Mahiul Islam +2
As powerful Large Language Models (LLMs) are now widely used for numerous practical applications, their safety is of critical importance. While alignment techniques have significan…