3 papers
cs.RO2026
Grounding Hierarchical Vision-Language-Action Models Through Explicit Language-Action Alignment
Theodor Wulff, Federico Tavella, Rahul Singh Maharjan +2
Achieving robot transparency is a critical step toward effective human-robot collaboration. To be transparent, a robot's natural language communication must be consistent with its…
cs.RO2025
Joint Action Language Modelling for Transparent Policy Execution
Theodor Wulff, Rahul Singh Maharjan, Xinyun Chi +1
An agent's intention often remains hidden behind the black-box nature of embodied policies. Communication using natural language statements that describe the next action can provid…
cs.CV2025
Balancing long- and short-term dynamics for the modeling of saliency in videos
Theodor Wulff, Fares Abawi, Philipp Allgeuer +1
The role of long- and short-term dynamics towards salient object detection in videos is under-researched. We present a Transformer-based approach to learn a joint representation of…