Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
Benchmarking Large Language Models with Integer Sequence Generation Tasks
Daniel O'Malley, Manish Bhattarai, Nishath Rajiv Ranasinghe +2
We present a novel benchmark designed to rigorously evaluate the capabilities of large language models (LLMs) in mathematical reasoning and algorithmic code synthesis tasks. The be…
cs.LG2025
The Transparent Earth: A Multimodal Foundation Model for the Earth's Subsurface
Arnab Mazumder, Javier E. Santos, Noah Hobbs +2
We present the Transparent Earth, a transformer-based architecture for reconstructing subsurface properties from heterogeneous datasets that vary in sparsity, resolution, and modal…