2 papers
cs.AR2025
Sangam: Chiplet-Based DRAM-PIM Accelerator with CXL Integration for LLM Inferencing
Khyati Kiyawat, Zhenxing Fan, Yasas Seneviratne +5
Large Language Models (LLMs) are becoming increasingly data-intensive due to growing model sizes, and they are becoming memory-bound as the context length and, consequently, the ke…
cs.AR2024
Swift: A Multi-FPGA Framework for Scaling Up Accelerated Graph Analytics
Oluwole Jaiyeoba, Abdullah T. Mughrabi, Morteza Baradaran +2
Graph analytics are vital in fields such as social networks, biomedical research, and graph neural networks (GNNs). However, traditional CPUs and GPUs struggle with the memory bott…