2 papers
cs.LG2026
annbatch unlocks terabyte-scale training of biological data in anndata
Ilan Gold, Felix Fischer, Lucas Arnoldt +2
The scale of biological datasets now routinely exceeds system memory, making data access rather than model computation the primary bottleneck in training machine-learning models. T…
q-bio.GN2026
GPU-accelerated single-cell analysis at scale with rapids-singlecell
Severin Dicks, Lukas Heumos, Lilly May +10
Single-cell sequencing technologies reveal cellular heterogeneity at high resolution, advancing our understanding of biological complexity. As datasets start to scale to tens of mi…