Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Fast Training Dataset Attribution via In-Context Learning
Milad Fotouhi, Mohammad Taha Bahadori, Oluwaseyi Feyisetan +2
We investigate the use of in-context learning and prompt engineering to estimate the contributions of training data in the outputs of instruction-tuned large language models (LLMs)…
cs.CL2024
Removing Spurious Correlation from Neural Network Interpretations
Milad Fotouhi, Mohammad Taha Bahadori, Oluwaseyi Feyisetan +2
The existing algorithms for identification of neurons responsible for undesired and harmful behaviors do not consider the effects of confounders such as topic of the conversation.…