3 papers
cs.RO2025
DiAReL: Reinforcement Learning with Disturbance Awareness for Robust Sim2Real Policy Transfer in Robot Control
Mohammadhossein Malmir, Josip Josifovski, Noah Klarmann +1
Delayed Markov decision processes (DMDPs) fulfill the Markov property by augmenting the state space of agents with a finite time window of recently committed actions. In reliance o…
cs.CV2025
Barlow-Swin: Toward a novel siamese-based segmentation architecture using Swin-Transformers
Morteza Kiani Haftlang, Mohammadhossein Malmir, Foroutan Parand +2
Medical image segmentation is a critical task in clinical workflows, particularly for the detection and delineation of pathological regions. While convolutional architectures like…
cs.RO2025
Safe Continual Domain Adaptation after Sim2Real Transfer of Reinforcement Learning Policies in Robotics
Josip Josifovski, Shangding Gu, Mohammadhossein Malmir +5
Domain randomization has emerged as a fundamental technique in reinforcement learning (RL) to facilitate the transfer of policies from simulation to real-world robotic applications…