3 papers
cs.LG2026
Model-Agnostic Solutions for Deep Reinforcement Learning in Non-Ergodic Contexts
Bert Verbruggen, Arne Vanhoyweghen, Vincent Ginis
Reinforcement Learning (RL) remains a central optimisation framework in machine learning. Although RL agents can converge to optimal solutions, the definition of ``optimality'' dep…
econ.GN2025
Metro 3 in Brussels under uncertainty: scenario-based public transport accessibility analysis
Brecht Verbeken, Arne Vanhoyweghen, Vincent Ginis
Metro Line 3 in Brussels is one of Europe's most debated infrastructure projects, marked by escalating costs, delays, and uncertainty over completion. Yet no public accessibility a…
cs.CL2025
Lexical Hints of Accuracy in LLM Reasoning Chains
Arne Vanhoyweghen, Brecht Verbeken, Andres Algaba +1
Fine-tuning Large Language Models (LLMs) with reinforcement learning to produce an explicit Chain-of-Thought (CoT) before answering produces models that consistently raise overall…