2 papers
cs.AR2026
C2P-Cache: Scalable GPU L1 Cache Sharing via Concurrent Candidate Pruning
Hanqing Li, Lizhou Wu, Tiejun Li +6
Modern GPUs rely on private per-SM L1 caches and a shared L2 cache, but this organization obscures cross-SM reuse: an L1 miss is typically forwarded to L2 even when the requested l…
cs.AR2025
NeuroPDE: A Neuromorphic PDE Solver Based on Spintronic and Ferroelectric Devices
Siqing Fu, Lizhou Wu, Tiejun Li +5
In recent years, new methods for solving partial differential equations (PDEs) such as Monte Carlo random walk methods have gained considerable attention. However, due to the lack…