3 papers
cs.LG2026
Hidden Gauge Controls Feature Specialization in ReLU Networks
Tongxi Wang
Training changes a network's predictions while allocating task-relevant structure across its internal units. In an overparameterized ReLU network, several neurons can begin with ex…
cs.AI2026
FBS: Modeling Native Parallel Reading inside a Transformer
Tongxi Wang
Large language models (LLMs) excel across many tasks, yet inference is still dominated by strictly token-by-token autoregression. Existing acceleration methods largely patch this p…
cs.LG2025
Stability of In-Context Learning: A Spectral Coverage Perspective
Tongxi Wang, Zhuoyang Xia
In-context learning (ICL) is a pivotal capability for the practical deployment of large-scale language models, yet its reliability can vary substantially with the number of demonst…