Showing 2026Show all
3 papers · 1 filter
eess.SY2026
Stochastic MPC with Online-optimized Policies and Closed-loop Guarantees
Marcell Bartos, Alexandre Didier, Jerome Sieber +2
This paper proposes a stochastic model predictive control method for linear systems affected by additive Gaussian disturbances that optimizes over disturbance feedback matrices onl…
cs.LG2026
A Predictive Law for On-Policy Self-Distillation From World Feedback
Tommy He, Jerome Sieber, Matteo Saponati
Moving beyond simple scalar rewards toward richer world feedback is a natural path to more scalable RL post-training. On-policy self-distillation (OPSD) is a promising recent appro…
cs.LG2026
Design Principles for Sequence Models via Coefficient Dynamics
Jerome Sieber, Antonio Orvieto, Melanie N. Zeilinger +1
Deep sequence models, ranging from Transformers and State Space Models (SSMs) to more recent approaches such as gated linear RNNs, fundamentally compute outputs as linear combinati…