Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
The Atlas of In-Context Learning: How Attention Heads Shape In-Context Retrieval Augmentation
Patrick Kahardipraja, Reduan Achtibat, Thomas Wiegand +2
Large language models are able to exploit in-context learning to access external knowledge beyond their training data through retrieval-augmentation. While promising, its inner wor…
cs.CL2025
A Close Look at Decomposition-based XAI-Methods for Transformer Language Models
Leila Arras, Bruno Puri, Patrick Kahardipraja +2
Various XAI attribution methods have been recently proposed for the transformer architecture, allowing for insights into the decision-making process of large language models by ass…