2 papers
cs.LG2026
PreLoRA: Hybrid Pre-training of Vision Transformers with Full Training and Low-Rank Adapters
Krishu K Thapa, Reet Barik, Krishna Teja Chitty-Venkata +2
Training large models ranging from millions to billions of parameters is highly resource-intensive, requiring significant time, compute, and memory. It is observed that most of the…
cs.LG2025
ForeSWE: Forecasting Snow-Water Equivalent with an Uncertainty-Aware Attention Model
Krishu K Thapa, Supriya Savalkar, Bhupinderjeet Singh +3
Various complex water management decisions are made in snow-dominant watersheds with the knowledge of Snow-Water Equivalent (SWE) -- a key measure widely used to estimate the water…