2 papers
cs.LG2026
Dispersion Loss Counteracts Embedding Condensation and Improves Generalization in Small Language Models
Chen Liu, Xingzhi Sun, Xi Xiao +8
Large language models (LLMs) achieve remarkable performance through ever-increasing parameter counts, but scaling incurs steep computational costs. To better understand LLM scaling…
cs.LG2025
STAGED: A Multi-Agent Neural Network for Learning Cellular Interaction Dynamics
Joao F. Rocha, Ke Xu, Xingzhi Sun +6
The advent of single-cell technology has significantly improved our understanding of cellular states and subpopulations in various tissues under normal and diseased conditions by e…