2 papers
cs.LG2026
Moonwalk: Inverse-Forward Differentiation
Dmitrii Krylov, Armin Karamzade, Roy Fox
Backpropagation's main limitation is its need to store intermediate activations (residuals) during the forward pass, which restricts the depth of trainable networks. This raises a…
cs.LG2024
Align Your Intents: Offline Imitation Learning via Optimal Transport
Maksim Bobrin, Nazar Buzun, Dmitrii Krylov +1
Offline Reinforcement Learning (RL) addresses the problem of sequential decision-making by learning optimal policy through pre-collected data, without interacting with the environm…