3 papers
cs.DC2026
DualDecoder: Accelerate Long Context LLM Inference by Predictive Prefetch
Zuning Liang, Zhiyi Yao, Qi Chen +6
Long-context inference is becoming a fundamental capability for modern LLM serving, especially driven by emerging agentic applications. Yet it faces a severe memory wall that the K…
cs.SE2026
FuzzPilot: Plateau-Triggered Recipe Validation for Structured Text Fuzzing
Zhiyi Yao
FuzzPilot is a controller for AFL++ that moves expensive reasoning out of the mutation hot path. When coverage plateaus, it snapshots the corpus, prepares candidate mutation recipe…
cs.DB2023
Minerva: Decentralized Collaborative Query Processing over InterPlanetary File System
Zhiyi Yao, Bowen Ding, Qianlan Bai +1
Data silos create barriers in accessing and utilizing data dispersed over networks. Directly sharing data easily suffers from the long downloading time, the single point failure an…