2 papers
eess.SY2025
TubeDAgger: Reducing the Number of Expert Interventions with Stochastic Reach-Tubes
Julian Lemmel, Manuel Kranzl, Adam Lamine +3
Interactive Imitation Learning deals with training a novice policy from expert demonstrations in an online fashion. The established DAgger algorithm trains a robust novice policy b…
cs.CE2025
Online Fine-Tuning of Carbon Emission Predictions using Real-Time Recurrent Learning for State Space Models
Julian Lemmel, Manuel Kranzl, Adam Lamine +3
This paper introduces a new approach for fine-tuning the predictions of structured state space models (SSMs) at inference time using real-time recurrent learning. While SSMs are kn…