3 papers
cs.DC2025
Memory Efficient and Staleness Free Pipeline Parallel DNN Training Framework with Improved Convergence Speed
Ankita Dutta, Nabendu Chaki, Rajat K. De
High resource requirement for Deep Neural Network (DNN) training across multiple GPUs necessitates development of various parallelism techniques. In this paper, we introduce two in…
cs.CL2025
CARE: A QLoRA-Fine Tuned Multi-Domain Chatbot With Fast Learning On Minimal Hardware
Ankit Dutta, Nabarup Ghosh, Ankush Chatterjee
Large Language models have demonstrated excellent domain-specific question-answering capabilities when finetuned with a particular dataset of that specific domain. However, fine-tu…
cs.DC2024
TiMePReSt: Time and Memory Efficient Pipeline Parallel DNN Training with Removed Staleness
Ankita Dutta, Nabendu Chaki, Rajat K. De
DNN training is time-consuming and requires efficient multi-accelerator parallelization, where a single training iteration is split over available accelerators. Current approaches…