1 paper
Amogh Sheth, Biruk Assefa, Yi Wen Huang +2
Large language models (LLMs) excel at multi-step reasoning but incur substantial inference cost. We introduce Causal Attribution Pruning (CAP), a training-free method that identifi…