collaborators

5 papers

cs.AI2026

Small Initialization Matters for Large Language Models

Liangkai Hang, Junjie Yao, Zhiyu Li +3

Large language models provide a tractable system for asking how intelligence itself emerges, rather than only how LLMs can be engineered. Although progress is usually attributed to…

cs.LG2025

Probability Signature: Bridging Data Semantics and Embedding Structure in Language Models

Junjie Yao, Zhi-Qin John Xu

The embedding space of language models is widely believed to capture the semantic relationships; for instance, embeddings of digits often exhibit an ordered structure that correspo…

math.NA2025

Solving multiscale dynamical systems by deep learning

Junjie Yao, Yuxiao Yi, Liangkai Hang +5

Multiscale dynamical systems, modeled by high-dimensional stiff ordinary differential equations (ODEs) with wide-ranging characteristic timescales, arise across diverse fields of s…

cs.LG2025

Scalable Complexity Control Facilitates Reasoning Ability of LLMs

Liangkai Hang, Junjie Yao, Zhiwei Bai +17

The reasoning ability of large language models (LLMs) has been rapidly advancing in recent years, attracting interest in more fundamental approaches that can reliably enhance their…

cs.CL2025

An Analysis for Reasoning Bias of Language Models with Small Initialization

Junjie Yao, Zhongwang Zhang, Zhi-Qin John Xu

Transformer-based Large Language Models (LLMs) have revolutionized Natural Language Processing by demonstrating exceptional performance across diverse tasks. This study investigate…