3 papers
cs.AR2026
LOCALUT: Harnessing Capacity-Computation Tradeoffs for LUT-Based Inference in DRAM-PIM
Junguk Hong, Changmin Shin, Sukjin Kim +6
Lookup tables (LUTs) have recently gained attention as an alternative compute mechanism that maps input operands to precomputed results, eliminating the need for arithmetic logic.…
cs.DC2026
DFLOP: A Data-driven Framework for Multimodal LLM Training Pipeline Optimization
Hyeonjun An, Sihyun Kim, Chaerim Lim +9
Multimodal Large Language Models (MLLMs) have achieved remarkable advances by integrating text, image, and audio understanding within a unified architecture. However, existing dist…
cs.LG2024
GraNNDis: Efficient Unified Distributed Training Framework for Deep GNNs on Large Clusters
Jaeyong Song, Hongsun Jang, Jaewon Jung +2
Graph neural networks (GNNs) are one of the rapidly growing fields within deep learning. While many distributed GNN training frameworks have been proposed to increase the training…