4k citations
- Nvidia (United States)US31 papers
- University of TorontoCA13 papers
- Stanford UniversityUS8 papers
- Massachusetts Institute of TechnologyUS6 papers
- University of California, MercedUS6 papers
- University of WashingtonUS6 papers
- Google (United States)US5 papers
- Seattle UniversityUS5 papers
- Technical University of MunichDE5 papers
- Arizona State UniversityUS4 papers
- Duke UniversityUS4 papers
- Georgia Institute of TechnologyUS4 papers
Showing 2026 · cs.DCShow all
2 papers · 2 filters
cs.DC2026
Fast Recovery for LLM Serving via Decoupled Device Memory Lifetime in Dynamo
Schwinn Saereesitthipitak, Mohammed Abdulwahhab, Hannah Zhang +6
Large language model (LLM) inference replicas run across tightly coupled GPUs and serve traffic continuously for weeks. Hardware and software failures are therefore inevitable, and…
cs.DC2026★ 3 cited
ALPHA-PIM: Analysis of Linear Algebraic Processing for High-Performance Graph Applications on a Real Processing-In-Memory System
Marzieh Barkhordar, Alireza Tabatabaeian, Mohammad Sadrosadati +5
Processing large-scale graph datasets is computationally intensive and time-consuming. Processor-centric CPU and GPU architectures, commonly used for graph applications, often face…