Publications (25)
Two Effects, One Trigger: On the Modality Gap, Object Bias, and Information Imbalance in Contrastive Vision-Language Models
Simon Schrodi, David T. Hoffmann, Max Argus +2
Contrastive vision-language models (VLMs), like CLIP, have gained popularity for their versatile applicability to various downstream tasks. Despite their successes in some tasks, l…
FreiHAND: A Dataset for Markerless Capture of Hand Pose and Shape from Single RGB Images
Christian Zimmermann, Duygu Ceylan, Jimei Yang +3
Estimating 3D hand pose from single RGB images is a highly ambiguous problem that relies on an unbiased training dataset. In this paper, we analyze cross-dataset generalization whe…
Compositional Servoing by Recombining Demonstrations
Max Argus, Abhijeet Nayak, Martin Büchner +3
Learning-based manipulation policies from image inputs often show weak task transfer capabilities. In contrast, visual servoing methods allow efficient task transfer in high-precis…
Efficient Learning of Object Placement with Intra-Category Transfer
Adrian Röfer, Russell Buchanan, Max Argus +2
Efficient learning from demonstration for long-horizon tasks remains an open challenge in robotics. While significant effort has been directed toward learning trajectories, a recen…
Concept Bottleneck Models Without Predefined Concepts
Simon Schrodi, Julian Schur, Max Argus +1
There has been considerable recent interest in interpretable concept-based models such as Concept Bottleneck Models (CBMs), which first predict human-interpretable concepts and the…
RobotIO: A Python Library for Robot Manipulation Experiments
Lukas Hermann, Max Argus, Adrian Roefer +2
Setting up robot environments to quickly test newly developed algorithms is still a difficult and time consuming process. This presents a significant hurdle to researchers interest…