49 citations · 49 across the 6 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
-MSHC: Unmasking Minimally Sufficient Head Circuits in Large Language Models with Experiments on Syntactic Classification Tasks
Pratim Chowdhary, Peter Chin, Deepernab Chakrabarty
Understanding which neural components drive specific capabilities in mid-sized language models (10B parameters) remains a key challenge. We introduce the -Minimu…
cs.CL2023
Adversarial Transformer Language Models for Contextual Commonsense Inference
Pedro Colon-Hernandez, Henry Lieberman, Yida Xin +3
Contextualized or discourse aware commonsense inference is the task of generating coherent commonsense assertions (i.e., facts) from a given story, and a particular sentence from t…