3 papers
cs.LG2026
Calibrated Partial Resets: Preventing Policy Collapse in Continual Reinforcement Learning
Luc McCutcheon, Evangelos Chatzaroulas, Saber Fallah
Neural networks are hindered by accumulating dormant neurons and loss of expressivity throughout training, particularly in non-stationary data settings, such as continual supervise…
cs.AI2025
Meta-World+: An Improved, Standardized, RL Benchmark
Reginald McLean, Evangelos Chatzaroulas, Luc McCutcheon +9
Meta-World is widely used for evaluating multi-task and meta-reinforcement learning agents, which are challenged to master diverse skills simultaneously. Since its introduction how…
cs.AI2024
In-Context Ensemble Learning from Pseudo Labels Improves Video-Language Models for Low-Level Workflow Understanding
Moucheng Xu, Evangelos Chatzaroulas, Luc McCutcheon +4
A Standard Operating Procedure (SOP) defines a low-level, step-by-step written guide for a business software workflow. SOP generation is a crucial step towards automating end-to-en…