2 papers
cs.LG2025
Pruning as a Defense: Reducing Memorization in Large Language Models
Mansi Gupta, Nikhar Waghela, Sarthak Gupta +2
Large language models have been shown to memorize significant portions of their training data, which they can reproduce when appropriately prompted. This work investigates the impa…
cs.LG2024
Confidence Is All You Need for MI Attacks
Abhishek Sinha, Himanshi Tibrewal, Mansi Gupta +2
In this evolving era of machine learning security, membership inference attacks have emerged as a potent threat to the confidentiality of sensitive data. In this attack, adversarie…