1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.AR2026
CIM-Tuner: Balancing the Compute and Storage Capacity of SRAM-CIM Accelerator via Hardware-mapping Co-exploration
Jinwu Chen, Yuhui Shi, He Wang +4
As an emerging type of AI computing accelerator, SRAM Computing-In-Memory (CIM) accelerators feature high energy efficiency and throughput. However, various CIM designs and under-e…
cs.AI2025
cuPilot: A Strategy-Coordinated Multi-agent Framework for CUDA Kernel Evolution
Jinwu Chen, Qidie Wu, Bin Li +5
Optimizing CUDA kernels is a challenging and labor-intensive task, given the need for hardware-software co-design expertise and the proprietary nature of high-performance kernel li…
cs.LG2024★ 1 cited
ScaleFold: Reducing AlphaFold Initial Training Time to 10 Hours
Feiwen Zhu, Arkadiusz Nowaczynski, Rundong Li +6
AlphaFold2 has been hailed as a breakthrough in protein folding. It can rapidly predict protein structures with lab-grade accuracy. However, its implementation does not include the…