From the 1 of 6 linked papers with an AI index.
6 papers
Hidden Gauge Controls Feature Specialization in ReLU Networks
Tongxi Wang
Training changes a network's predictions while allocating task-relevant structure across its internal units. In an overparameterized ReLU network, several neurons can begin with ex…
Tracking Drift: Variation-Aware Entropy Scheduling for Non-Stationary Reinforcement Learning
Tongxi Wang, Zhuoyang Xia, Xinran Chen +1
The paper proposes an adaptive method for adjusting the entropy coefficient in reinforcement learning to handle non‑stationary environments, using online drift proxies to scale exp…
Sharp Spectral Thresholds for Logit Fixed Points
Tongxi Wang
Softmax feedback systems are a common mathematical core of entropy-regularized reinforcement learning, logit game dynamics, population choice, and mean-field variational updates. T…
FBS: Modeling Native Parallel Reading inside a Transformer
Tongxi Wang
Large language models (LLMs) excel across many tasks, yet inference is still dominated by strictly token-by-token autoregression. Existing acceleration methods largely patch this p…
Stability of In-Context Learning: A Spectral Coverage Perspective
Tongxi Wang, Zhuoyang Xia
In-context learning (ICL) is a pivotal capability for the practical deployment of large-scale language models, yet its reliability can vary substantially with the number of demonst…
ChiEngMixBench: Evaluating Large Language Models on Expert-Style Chinese-English Terminology Mixing
Qingyan Yang, Tongxi Wang, Yunsheng Luo
Large language models increasingly mediate multilingual professional communication, where useful generation requires adapting to community conventions about which expressions are r…