2 papers
cs.AR2026
TOM: A Ternary Read-only Memory Accelerator for LLM-powered Edge Intelligence
Hongyi Guan, Yijia Zhang, Wenqiang Wang +4
The deployment of Large Language Models (LLMs) for real-time intelligence on edge devices is rapidly growing. However, conventional hardware architectures face a fundamental memory…
cs.AR2025
Efficient Architecture for RISC-V Vector Memory Access
Hongyi Guan, Yichuan Gao, Chenlu Miao +4
Vector processors frequently suffer from inefficient memory accesses, particularly for strided and segment patterns. While coalescing strided accesses is a natural solution, effect…