2 papers
cs.CL2026
Word Recovery in Large Language Models Enables Character-Level Tokenization Robustness
Zhipeng Yang, Shu Yang, Lijie Hu +1
Large language models (LLMs) trained with canonical tokenization exhibit surprising robustness to non-canonical inputs such as character-level tokenization, yet the mechanisms unde…
cs.CL2025
Internal Chain-of-Thought: Empirical Evidence for Layer-wise Subtask Scheduling in LLMs
Zhipeng Yang, Junzhuo Li, Siyu Xia +1
We show that large language models (LLMs) exhibit an : they sequentially decompose and execute composite tasks layer-by-layer. Two claims ground…