papers
Publications (10)
cs.AI2026
Data-driven Circuit Discovery for Interpretability of Language Models
Daking Rai, Mor Geva, Ziyu Yao
cs.LG2025
A Survey on Sparse Autoencoders: Interpreting the Internal Mechanisms of Large Language Models
Dong Shu, Xuansheng Wu, Haiyan Zhao +4
cs.AI2024
An Investigation of Neuron Activation as a Unified Lens to Explain Chain-of-Thought Eliciting Arithmetic Reasoning of LLMs
Daking Rai, Ziyu Yao
cs.AI2025
A Practical Review of Mechanistic Interpretability for Transformer-Based Language Models
Daking Rai, Yilun Zhou, Shi Feng +2
cs.CL2023
Improving Generalization in Language Model-Based Text-to-SQL Semantic Parsing: Two Simple Semantic Boundary-Based Techniques
Daking Rai, Bailin Wang, Yilun Zhou +1
cs.SE2025
Mechanistic Understanding of Language Models in Syntactic Code Completion
Samuel Miller, Daking Rai, Ziyu Yao
cs.IR2024
Understanding the Effect of Algorithm Transparency of Model Explanations in Text-to-SQL Semantic Parsing
Daking Rai, Rydia R. Weiland, Kayla Margaret Gabriella Herrera +2
cs.CL2025
All for One: LLMs Solve Mental Math at the Last Token With Information Transferred From Other Tokens
Siddarth Mamidanna, Daking Rai, Ziyu Yao +1
cs.CL2026
Failure by Interference: Language Models Make Balanced Parentheses Errors When Faulty Mechanisms Overshadow Sound Ones
Daking Rai, Samuel Miller, Kevin Moran +1
cs.CL2023
Explaining Large Language Model-Based Neural Semantic Parsers (Student Abstract)
Daking Rai, Yilun Zhou, Bailin Wang +1