2 papers
cs.CL2023
Counterfactually Probing Language Identity in Multilingual Models
Anirudh Srinivasan, Venkata S Govindarajan, Kyle Mahowald
Techniques in causal analysis of language models illuminate how linguistic information is organized in LLMs. We use one such technique, AlterRep, a method of counterfactual probing…
cs.CL2022
longhorns at DADC 2022: How many linguists does it take to fool a Question Answering model? A systematic approach to adversarial attacks
Venelin Kovatchev, Trina Chatterjee, Venkata S Govindarajan +9
Developing methods to adversarially challenge NLP systems is a promising avenue for improving both model performance and interpretability. Here, we describe the approach of the tea…