2 papers
cs.CR2026
Testing LLM Arithmetic Reasoning Generalization with Automatic Numeric-Remapping Attacks
Malia Barker, Bishal Lakha, Edoardo Serra +1
Large language models achieve strong performance on arithmetic reasoning benchmarks, and one common response to arithmetic brittleness is to delegate computation to code. Yet model…
cs.LG2026
Scaling Novel Graph Generation via Lightweight Structure-Guided Autoregressive Models
Alessio Barboni, Massimiliano Lupo Pasini, Bishal Lakha +1
Generating realistic and diverse graphs is a key problem in machine learning, with applications in molecular discovery, circuit design, cybersecurity, and beyond. However, current…