collaborators

6 papers

cs.CL2026

Measuring Maximum Activations in Open Large Language Models

Luxuan Chen, Han Tian, Xinran Chen +9

The dynamic range of activations is a first-order constraint for low-bit quantization, activation scaling, and stable LLM inference. Prior work characterized outlier features and m…

cs.CL2026

EndPrompt: Efficient Long-Context Extension via Terminal Anchoring

Han Tian, Luxuan Chen, Xinran Chen +10

Extending the context window of large language models typically requires training on sequences at the target length, incurring quadratic memory and computational costs that make lo…

cs.CL2026

Towards AI Search Paradigm

Yuchen Li, Hengyi Cai, Rui Kong +20

In this paper, we introduce the AI Search Paradigm, a comprehensive blueprint for next-generation search systems capable of emulating human information processing and decision-maki…

cs.DC2026

FlexSpec: Frozen Drafts Meet Evolving Targets in Edge-Cloud Collaborative LLM Speculative Decoding

Yuchen Li, Rui Kong, Zhonghao Lyu +11

Deploying large language models (LLMs) in mobile and edge computing environments is constrained by limited on-device resources, scarce wireless bandwidth, and frequent model evolut…

cs.CV2025

VPN: Visual Prompt Navigation

Shuo Feng, Zihan Wang, Yuchen Li +6

While natural language is commonly used to guide embodied agents, the inherent ambiguity and verbosity of language often hinder the effectiveness of language-guided navigation in c…

cs.IR2025

AI Guided Accelerator For Search Experience

Jayanth Yetukuri, Mehran Elyasi, Samarth Agrawal +4

Effective query reformulation is pivotal in narrowing the gap between a user's exploratory search behavior and the identification of relevant products in e-commerce environments. W…