2 papers
eess.SY2026
An Output Feedback Q-learning Algorithm for Optimal Control of Nonlinear Systems with Koopman Linear Embedding
Victor G. Lopez, Malte Heinrich, Matthias A. Müller
In the reinforcement learning literature, strong theoretical guarantees have been obtained for algorithms applicable to LTI systems. However, in the nonlinear case only weaker resu…
cs.LG2026
Tuning the burn-in phase in training recurrent neural networks improves their performance
Julian D. Schiller, Malte Heinrich, Victor G. Lopez +1
Training recurrent neural networks (RNNs) with standard backpropagation through time (BPTT) can be challenging, especially in the presence of long input sequences. A practical alte…