3 papers
cs.LG2025
EAP-GP: Mitigating Saturation Effect in Gradient-based Automated Circuit Identification
Lin Zhang, Wenshuo Dong, Zhuoran Zhang +5
Understanding the internal mechanisms of transformer-based language models remains challenging. Mechanistic interpretability based on circuit discovery aims to reverse engineer neu…
cs.LG2025
Evaluating Data Influence in Meta Learning
Chenyang Ren, Huanyi Xie, Shu Yang +3
As one of the most fundamental models, meta learning aims to effectively address few-shot learning challenges. However, it still faces significant issues related to the training da…
cs.AI2024
Understanding Reasoning in Chain-of-Thought from the Hopfieldian View
Lijie Hu, Liang Liu, Shu Yang +5
Large Language Models have demonstrated remarkable abilities across various tasks, with Chain-of-Thought (CoT) prompting emerging as a key technique to enhance reasoning capabiliti…