3 papers
cs.CL2025
Automated MCQA Benchmarking at Scale: Evaluating Reasoning Traces as Retrieval Sources for Domain Adaptation of Small Language Models
Ozan Gokdemir, Neil Getty, Robert Underwood +5
As scientific knowledge grows at an unprecedented pace, evaluation benchmarks must evolve to reflect new discoveries and ensure language models are tested on current, diverse liter…
cs.LG2025
Benchmarking community drug response prediction models: datasets, models, tools, and metrics for cross-dataset generalization analysis
Alexander Partin, Priyanka Vasanthakumari, Oleksandr Narykov +17
Deep learning (DL) and machine learning (ML) models have shown promise in drug response prediction (DRP), yet their ability to generalize across datasets remains an open question,…
q-bio.BM2024
Assessing Reusability of Deep Learning-Based Monotherapy Drug Response Prediction Models Trained with Omics Data
Jamie C. Overbeek, Alexander Partin, Thomas S. Brettin +21
Cancer drug response prediction (DRP) models present a promising approach towards precision oncology, tailoring treatments to individual patient profiles. While deep learning (DL)…