2 papers
cs.LG2026
Radial Suppression Accelerates Algorithmic Generalization: A Geometric Analysis of Delayed Generalization
Srijan Tiwari, Aditya Chauhan, Manjot Singh
Why do neural networks memorize algorithmic training data long before they generalize? We present a geometric case study demonstrating that, on tasks where generalization requires…
cs.LG2026
Structural Disentanglement in Bilinear MLPs via Architectural Inductive Bias
Ojasva Nema, Kaustubh Sharma, Aditya Chauhan +1
Selective unlearning and long-horizon extrapolation remain fragile in modern neural networks, even when tasks have underlying algebraic structure. In this work, we argue that these…