Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Mechanistic Analysis of Alignment Algorithms in Language Models
Aarush Sinha, Ishan Garg, Veeraraju Elluru +2
Post-training alignment algorithms are predominantly evaluated as black boxes, obscuring how they reshape language models' internal computations. We present a systematic mechanisti…
cs.LG2024
A Multi-Branched Radial Basis Network Approach to Predicting Complex Chaotic Behaviours
Aarush Sinha
In this study, we propose a multi branched network approach to predict the dynamics of a physics attractor characterized by intricate and chaotic behavior. We introduce a unique ne…