collaborators

6 papers

cs.CL2026

Measuring Maximum Activations in Open Large Language Models

Luxuan Chen, Han Tian, Xinran Chen +9

The dynamic range of activations is a first-order constraint for low-bit quantization, activation scaling, and stable LLM inference. Prior work characterized outlier features and m…

cs.CL2026

EndPrompt: Efficient Long-Context Extension via Terminal Anchoring

Han Tian, Luxuan Chen, Xinran Chen +10

Extending the context window of large language models typically requires training on sequences at the target length, incurring quadratic memory and computational costs that make lo…

cs.CL2026

Towards AI Search Paradigm

Yuchen Li, Hengyi Cai, Rui Kong +20

In this paper, we introduce the AI Search Paradigm, a comprehensive blueprint for next-generation search systems capable of emulating human information processing and decision-maki…

cs.CL2026

Evaluating LLM-based Agents for Multi-Turn Conversations: A Survey

Shengyue Guan, Jindong Wang, Jiang Bian +3

This survey examines evaluation methods for large language model (LLM)-based agents in multi-turn conversational settings. Using a PRISMA-inspired framework, we systematically revi…

cs.DC2026

FlexSpec: Frozen Drafts Meet Evolving Targets in Edge-Cloud Collaborative LLM Speculative Decoding

Yuchen Li, Rui Kong, Zhonghao Lyu +11

Deploying large language models (LLMs) in mobile and edge computing environments is constrained by limited on-device resources, scarce wireless bandwidth, and frequent model evolut…

q-bio.BM2024

Pre-trained Molecular Language Models with Random Functional Group Masking

Tianhao Peng, Yuchen Li, Xuhong Li +7

Recent advancements in computational chemistry have leveraged the power of trans-former-based language models, such as MoLFormer, pre-trained using a vast amount of simplified mole…