1 paper
Prateek Kumar Sikdar, Arpan Ghosh
Layer-skipping methods for efficient LLM inference decide, at some granularity, which transformer layers to execute for a given input. We present a rigor-matched, three-seed audit…