3 papers
cs.CL2026
Gated Tree Cross-Attention for Checkpoint-Compatible Syntax Injection in Decoder-Only LLMs
Xinyu Gao, Shaonan Wang, Nai Ding
Decoder-only large language models achieve strong broad performance but are brittle to minor grammatical perturbations, undermining reliability for downstream reasoning. However, d…
cs.CL2026
Component-Level Lesioning of Language Models Reveals Clinically Aligned Aphasia Phenotypes
Yifan Wang, Jichen Zheng, Jingyuan Sun +5
Large language models (LLMs) increasingly exhibit human-like linguistic behaviors and internal representations that they could serve as computational simulators of language cogniti…
cs.CL2025
How Syntax Specialization Emerges in Language Models
Xufeng Duan, Zhaoqian Yao, Yunhao Zhang +2
Large language models (LLMs) have been found to develop surprising internal specializations: Individual neurons, attention heads, and circuits become selectively sensitive to synta…