3 papers
cs.AR2026
You Only Charge Once 2.0 : A End-to-End Analog Computing-in-Memory Architecture with Reconfigurable Switched Capacitors
Zihao Xuan, Yewen Li, Jia Chen +3
Analog Computing-in-Memory (ACiM) accelerates deep neural networks by keeping weights inside memory arrays and executing dot products in the analog domain. However, modern ACiM acc…
cs.AR2026
FusionCIM: Accelerating LLM Inference with Fusion-Driven Computing-in-Memory Architecture
Zihao Xuan, Jia Chen, Yewen Li +4
In this paper, we propose FusionCIM, an operator-fusion-driven compute-in-memory (CIM) accelerator architecture for efficient and scalable LLM inference, with three key innovations…
q-bio.GN2025
FastDup: a scalable duplicate marking tool using speculation-and-test mechanism
Zhonghai Zhang, Yewen Li, Ke Meng +2
Duplicate marking is a critical preprocessing step in gene sequence analysis to flag redundant reads arising from polymerase chain reaction(PCR) amplification and sequencing artifa…