3 papers
cs.AI2026
The Importance of Being Statistically Earnest: A Critical Re-evaluation of GSM-Symbolic
Dominika Agnieszka DÅugosz, Arlindo Oliveira, Natalia DÃaz-RodrÃguez
The GSM-Symbolic benchmark (Mirzadeh et al., 2025) reported consistent performance drops across 25 Large Language Models (LLMs) when tested on template-generated variants of GSM8K…
cs.LG2026
Engineering FAIR Privacy-preserving Applications that Learn Histories of Disease
Ines N. Duarte, Praphulla M. S. Bhawsar, Lee K. Mason +4
A recent report on "Learning the natural history of human disease with generative transformers" created an opportunity to assess the engineering challenge of delivering user-facing…
cs.LG2025
Fractal Language Modelling by Universal Sequence Maps (USM)
Jonas S Almeida, Daniel E Russ, Susana Vinga +6
Motivation: With the advent of Language Models using Transformers, popularized by ChatGPT, there is a renewed interest in exploring encoding procedures that numerically represent s…