4 papers
Performance of Small Language Model Pretraining on FABRIC: An Empirical Study
Praveen Rao
Large language models (LLMs) require enormous computing power to pretrain on massive datasets. When limited datasets are available, smaller-sized LLMs are better choice to pretrain…
Optimizing the Variant Calling Pipeline Execution on Human Genomes Using GPU-Enabled Machines
Ajay Kumar, Praveen Rao, Peter Sanders
Variant calling is the first step in analyzing a human genome and aims to detect variants in an individual's genome compared to a reference genome. Due to the computationally-inten…
Evaluating the Performance of AI Text Detectors, Few-Shot and Chain-of-Thought Prompting Using DeepSeek Generated Text
Hulayyil Alshammari, Praveen Rao
Large language models (LLMs) have rapidly transformed the creation of written materials. LLMs have led to questions about writing integrity, thereby driving the creation of artific…
Seventeenth-Century Spanish American Notary Records for Fine-Tuning Spanish Large Language Models
Shraboni Sarker, Ahmad Tamim Hamad, Hulayyil Alshammari +2
Large language models have gained tremendous popularity in domains such as e-commerce, finance, healthcare, and education. Fine-tuning is a common approach to customize an LLM on a…