3 papers
cs.CL2024
Adjusting Interpretable Dimensions in Embedding Space with Human Judgments
Katrin Erk, Marianna Apidianaki
Embedding spaces contain interpretable dimensions indicating gender, formality in style, or even object properties. This has been observed multiple times. Such interpretable dimens…
cs.CL2023
A Method for Studying Semantic Construal in Grammatical Constructions with Interpretable Contextual Embedding Spaces
Gabriella Chronis, Kyle Mahowald, Katrin Erk
We study semantic construal in grammatical constructions using large language models. First, we project contextual word embeddings into three interpretable semantic spaces, each de…
cs.CL2022
longhorns at DADC 2022: How many linguists does it take to fool a Question Answering model? A systematic approach to adversarial attacks
Venelin Kovatchev, Trina Chatterjee, Venkata S Govindarajan +9
Developing methods to adversarially challenge NLP systems is a promising avenue for improving both model performance and interpretability. Here, we describe the approach of the tea…