3 papers
cs.AI2026
RL Post-Training Builds Compositional Reasoning Strategies
Azwar Abdulsalam, Nishil Patel, Andrew Saxe
Does RL post-training merely amplify primitive skills already latent in a base model, or can it compose primitive skills into new higher-level strategies? We study this question in…
cs.LG2023
Learning Recurrent Models with Temporally Local Rules
Azwar Abdulsalam, Joseph G. Makin
Fitting generative models to sequential data typically involves two recursive computations through time, one forward and one backward. The latter could be a computation of the loss…
q-fin.PR2020
On the Pricing of Currency Options under Variance Gamma Process
Azwar Abdulsalam, Gowri Jayprakash, Abhijeet Chandra
The pricing of currency options is largely dependent on the dynamic relationship between a pair of currencies. Typically, the pricing of options with payoffs dependent on multi-ass…