2 papers
cs.LG2026
Radial Suppression Accelerates Algorithmic Generalization: A Geometric Analysis of Delayed Generalization
Srijan Tiwari, Aditya Chauhan, Manjot Singh
Why do neural networks memorize algorithmic training data long before they generalize? We present a geometric case study demonstrating that, on tasks where generalization requires…
cs.LG2026
Mechanistic Evidence for Spectral Structures in Prior-Data Fitted Networks
Kaustubh Sharma, Srijan Tiwari, Ojasva Nema +1
Prior-Data Fitted Networks (PFNs) enable amortized Bayesian inference in a single forward pass, yet their internal representations remain opaque. It is unknown whether PFNs encode…