Showing cs.LGShow all
3 papers · 1 filter
cs.LG2025
Binary Autoencoder for Mechanistic Interpretability of Large Language Models
Hakaze Cho, Haolin Yang, Yanshu Li +2
Existing works are dedicated to untangling atomized numerical components (features) from the hidden states of Large Language Models (LLMs). However, they typically rely on autoenco…
cs.LG2025
Pareto Optimal Algorithmic Recourse in Multi-cost Function
Wen-Ling Chen, Hong-Chang Huang, Kai-Hung Lin +2
In decision-making systems, algorithmic recourse aims to identify minimal-cost actions to alter an individual features, thereby obtaining a desired outcome. This empowers individua…
cs.LG2025
PXGen: A Post-hoc Explainable Method for Generative Models
Yen-Lung Huang, Ming-Hsi Weng, Hao-Tsung Yang
With the rapid growth of generative AI in numerous applications, explainable AI (XAI) plays a crucial role in ensuring the responsible development and deployment of generative AI t…