2 papers
cs.LG2026
Concept Component Analysis: A Principled Approach for Concept Extraction in LLMs
Yuhang Liu, Erdun Gao, Dong Gong +2
Developing human understandable interpretation of large language models (LLMs) becomes increasingly critical for their deployment in essential domains. Mechanistic interpretability…
cs.LG2025
Analytic DAG Constraints for Differentiable DAG Learning
Zhen Zhang, Ignavier Ng, Dong Gong +6
Recovering the underlying Directed Acyclic Graph (DAG) structures from observational data presents a formidable challenge, partly due to the combinatorial nature of the DAG-constra…