2 papers
cs.LG2026
Chebyshev Policies and the Mountain Car Problem: Reinforcement Learning for Low-Dimensional Control Tasks
Stefan Huber, Hannes Unger, Georg Schäfer +2
We analytically solve the Mountain Car problem, a canonical benchmark in RL, and derive an optimal control solution, closing a gap after 36 years. This enables us to reveal two sur…
cs.DB2026
Reconstructing OPC UA Address Spaces from Time-Series Databases
Lukas Lürzer, Hannes Unger, Stefan Huber
OPC UA has become the dominant open protocol in operational technology. Time-series databases routinely archive OPC UA telemetry but discard the semantic metadata (node hierarchy,…