3 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.CL2024★ 2 cited
Extracting Paragraphs from LLM Token Activations
Nicholas Pochinkov, Angelo Benoit, Lovkush Agarwal +2
Generative large language models (LLMs) excel in natural language processing tasks, yet their inner workings remain underexplored beyond token-level predictions. This study investi…
cs.LG2021★ 3 cited
On Locality of Local Explanation Models
Sahra Ghalebikesabi, Lucile Ter-Minassian, Karla Diaz-Ordaz +1
Shapley values provide model agnostic feature attributions for model outcome at a particular instance by simulating feature absence under a global population distribution. The use…