2 papers
cs.LG2026
Towards Poisoning Robustness Certification for Natural Language Generation
Mihnea Ghitu, Matthew Wicker
Understanding the reliability of natural language generation is critical for deploying foundation models in security-sensitive domains. While certified poisoning defenses provide p…
cs.LG2025
Model Guidance via Robust Feature Attribution
Mihnea Ghitu, Vihari Piratla, Matthew Wicker
Controlling the patterns a model learns is essential to preventing reliance on irrelevant or misleading features. Such reliance on irrelevant features, often called shortcut featur…