2 papers
cs.CL2026
MACRO: Markov Chain Routing of Transformer Layers
Paweł Batorski, Abtin Pourhadi, Akylgali Aitaza +2
Standard Large Language Models (LLMs) execute layers sequentially. Dynamic layer routing, i.e. search for a different execution path through layers involving layer repetitions, ski…
cs.LG2026
GROM: Gradient-Free Rapid One-Shot Machine Unlearning
Paweł Batorski, Przemysław Spurek, Paul Swoboda
Machine unlearning has become a critical capability for safely removing specific, sensitive knowledge from large language models (LLMs). Current state-of-the-art approaches primari…