2 papers
cs.CL2025
Memorization and Knowledge Injection in Gated LLMs
Xu Pan, Ely Hahami, Zechen Zhang +1
Large Language Models (LLMs) currently struggle to sequentially add new memories and integrate new knowledge. These limitations contrast with the human ability to continuously lear…
cs.LG2025
When narrower is better: the narrow width limit of Bayesian parallel branching neural networks
Zechen Zhang, Haim Sompolinsky
The infinite width limit of random neural networks is known to result in Neural Networks as Gaussian Process (NNGP) (Lee et al. (2018)), characterized by task-independent kernels.…