2 papers
cs.DC2026
SLO-Aware Compute Resource Allocation for Prefill-Decode Disaggregated LLM Inference
Luchang Li, Dongfang Li, Bozhao Gong +1
Prefill-Decode (P/D) disaggregation has emerged as a widely adopted optimization strategy for Large Language Model (LLM) inference. However, there currently exists no well-establis…
cs.LG2025
A Study on Regularization-Based Continual Learning Methods for Indic ASR
Gokul Adethya T, S. Jaya Nirmala
Indias linguistic diversity poses significant challenges for developing inclusive Automatic Speech Recognition (ASR) systems. Traditional multilingual models, which require simulta…