2 papers
cs.LG2026
Scalable Pretraining of Large Mixture of Experts Language Models on Aurora Super Computer
Dharma Teja Vooturi, Dhiraj Kalamkar, Dipankar Das +1
Pretraining Large Language Models (LLMs) from scratch requires massive amount of compute. Aurora super computer is an ExaScale machine with 127,488 Intel PVC (Ponte Vechio) GPU til…
astro-ph.SR2025
Multi-modal encoder-decoder neural network for forecasting solar wind speed at L1
Dattaraj B. Dhuri, Shravan M. Hanasoge, Harsh Joon +3
The solar wind, accelerated within the solar corona, sculpts the heliosphere and continuously interacts with planetary atmospheres. On Earth, high-speed solar-wind streams may lead…