3 papers
cs.LG2026
Performance of Small Language Model Pretraining on FABRIC: An Empirical Study
Praveen Rao
Large language models (LLMs) require enormous computing power to pretrain on massive datasets. When limited datasets are available, smaller-sized LLMs are better choice to pretrain…
cs.DC2025
Optimizing the Variant Calling Pipeline Execution on Human Genomes Using GPU-Enabled Machines
Ajay Kumar, Praveen Rao, Peter Sanders
Variant calling is the first step in analyzing a human genome and aims to detect variants in an individual's genome compared to a reference genome. Due to the computationally-inten…
cs.CL2025
Evaluating the Performance of AI Text Detectors, Few-Shot and Chain-of-Thought Prompting Using DeepSeek Generated Text
Hulayyil Alshammari, Praveen Rao
Large language models (LLMs) have rapidly transformed the creation of written materials. LLMs have led to questions about writing integrity, thereby driving the creation of artific…