1 paper
Zhenyu Cui, Xiangzhong Luo
Recent mechanistic studies suggest that large language models (LLMs) may utilize their depth inefficiently in standard single-turn tasks. Whether this still holds in autonomous age…