3 papers
cs.LG2026
SignMuon: Communication-Efficient Distributed Muon Optimization
Neel Mishra, Kushagara Trivedi, Pawan Kumar
Distributed training of large neural networks is bottlenecked by full-precision gradient communication and by coordinatewise optimizers that ignore the matrix structure of weight t…
cs.LG2025
Hierarchical Sparse Plus Low Rank Compression of LLM
Pawan Kumar, Aditi Gupta
Modern large language models (LLMs) place extraordinary pressure on memory and compute budgets, making principled compression indispensable for both deployment and continued traini…
cs.CV2025
A Fast and Efficient Modern BERT based Text-Conditioned Diffusion Model for Medical Image Segmentation
Venkata Siddharth Dhara, Pawan Kumar
In recent times, denoising diffusion probabilistic models (DPMs) have proven effective for medical image generation and denoising, and as representation learners for downstream seg…