3 papers
cs.LG2025
AdaptiveK: Complexity-Driven Sparse Autoencoders for Interpretable Language Model Representations
Yifei Yao, Hanrong Zhang, Mengnan Du
Understanding the internal representations of large language models (LLMs) remains a central challenge for interpretability research. Sparse autoencoders (SAEs) offer a promising s…
cs.AI2024
SAKA: An Intelligent Platform for Semi-automated Knowledge Graph Construction and Application
Hanrong Zhang, Xinyue Wang, Jiabao Pan +1
Knowledge graph (KG) technology is extensively utilized in many areas, and many companies offer applications based on KG. Nonetheless, most KG platforms necessitate expertise and t…
cs.CV2024
Invisible Backdoor Attack against Self-supervised Learning
Hanrong Zhang, Zhenting Wang, Boheng Li +7
Self-supervised learning (SSL) models are vulnerable to backdoor attacks. Existing backdoor attacks that are effective in SSL often involve noticeable triggers, like colored patche…