37 citations · 55 across the 8 of their papers we have counts for
8 papers
Distributed Cross-Channel Hierarchical Aggregation for Foundation Models
Aristeidis Tsaris, Isaac Lyngaas, John Lagregren +6
Vision-based scientific foundation models hold significant promise for advancing scientific discovery and innovation. This potential stems from their ability to aggregate images fr…
Adapting CSI-Guided Imaging Across Diverse Environments: An Experimental Study Leveraging Continuous Learning
Cheng Chen, Shoki Ohta, Takayuki Nishio +1
This study explores the feasibility of adapting CSI-guided imaging across varied environments. Focusing on continuous model learning through continuous updates, we investigate CSI-…
CG-Kit: Code Generation Toolkit for Performant and Maintainable Variants of Source Code Applied to Flash-X Hydrodynamics Simulations
Johann Rudi, Youngjun Lee, Aidan H. Chadha +4
CG-Kit is a new code generation toolkit that we propose as a solution for portability and maintainability for scientific computing applications. The development of CG-Kit is rooted…
Ultra-Long Sequence Distributed Transformer
Xiao Wang, Isaac Lyngaas, Aristeidis Tsaris +7
Transformer models trained on long sequences often achieve higher accuracy than short sequences. Unfortunately, conventional transformers struggle with long sequence training due t…
KAKURENBO: Adaptively Hiding Samples in Deep Neural Network Training
Truong Thao Nguyen, Balazs Gerofi, Edgar Josafat Martinez-Noriega +2
This paper proposes a method for hiding the least-important samples during the training of deep neural networks to increase efficiency, i.e., to reduce the cost of training. Using…
Exploiting Scratchpad Memory for Deep Temporal Blocking: A case study for 2D Jacobian 5-point iterative stencil kernel (j2d5pt)
Lingqi Zhang, Mohamed Wahib, Peng Chen +4
General Purpose Graphics Processing Units (GPGPU) are used in most of the top systems in HPC. The total capacity of scratchpad memory has increased by more than 40 times in the las…