activity
20222026
most citedEmpowering Edge Intelligence: A Comprehensive Survey on On-Device AI Models

185 citations · 214 across the 25 of their papers we have counts for

collaborators

29 papers

cs.SE2026

EchoFuzz: Empowering Smart Contract Fuzzing with Large Language Models

Juanen Li, Peng Qian, Guanyan Li +6

Smart contracts, serving as the cornerstone of decentralized applications, autonomously manage trillion-dollar digital assets, making them attractive targets for attacks. Fuzzing h…

cs.DS2026

Dense Weak Hiding: Closing Complexity Gaps in Nonconvex and PL Finite-Sum Optimization under Individual Smoothness

Yuxing Peng, Zhiqing Tang, Weijia Jia

Under individual smoothness, the optimal incremental first-order oracle (IFO) complexity of nonconvex finite-sum optimization is open. Known algorithms use $O(n+\sqrt n\,ΔL_{\max}/…

cs.CV2026

FeatFix: Reuse What You Verify through Local Exact-Feature Correction for Faster Cached Diffusion Inference

Hanshuai Cui, Zhiqing Tang, Zhi Yao +3

Diffusion models are widely used to generate high-quality images and videos, but their iterative denoising process remains computationally intensive. A growing class of training-fr…

cs.AI2026

MemTxn: A Transaction Boundary for Source-Supported Updates and Complete-State Recovery in Agent Memory

Hanshuai Cui, Zhiqing Tang, Zhi Yao +3

Persistent memory lets long-running large language model agents reuse information across sessions and tasks. Yet errors in writable memory can persist and corrupt future behavior.…

cs.DC2026

LASER: Load-Aware Serving with Early-Exit for Reasoning LLMs at the Edge

Zhiqing Tang, Size Li, Hanshuai Cui +5

Large reasoning models (LRMs) such as DeepSeek-R1 have achieved strong performance through extended chain-of-thought (CoT) generation. However, deploying them on edge devices raise…

cs.DC2026

RISE: Relay Inference and Online Scheduling for Efficient Edge-Device Collaborative Diffusion Model Services

Zilan Huang, Zhiqing Tang, Hanshuai Cui +4

Text-to-image diffusion models are increasingly deployed at the network edge to serve heterogeneous workloads with diverse quality and latency requirements. However, existing deplo…