Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
A Single Layer to Explain Them All:Understanding Massive Activations in Large Language Models
Zeru Shi, Zhenting Wang, Fan Yang +2
We investigate the origins of massive activations in large language models (LLMs) and identify a specific layer named the \textbf{Massive Emergence Layer (ME Layer)}, that is consi…
cs.CL2024
Uncertainty is Fragile: Manipulating Uncertainty in Large Language Models
Qingcheng Zeng, Mingyu Jin, Qinkai Yu +12
Large Language Models (LLMs) are employed across various high-stakes domains, where the reliability of their outputs is crucial. One commonly used method to assess the reliability…