2 citations · 3 across the 3 of their papers we have counts for
Showing cs.DCShow all
2 papers · 1 filter
cs.DC2025★ 1 cited
Efficient LLM Inference with Activation Checkpointing and Hybrid Caching
Sanghyeon Lee, Hongbeen Kim, Soojin Hwang +3
Recent large language models (LLMs) with enormous model sizes use many GPUs to meet memory capacity requirements incurring substantial costs for token generation. To provide cost-e…
cs.DC2021★ 2 cited
Hardware-assisted Trusted Memory Disaggregation for Secure Far Memory
Taekyung Heo, Seunghyo Kang, Sanghyeon Lee +2
Memory disaggregation provides efficient memory utilization across network-connected systems. It allows a node to use part of memory in remote nodes in the same cluster. Recent stu…