From the 1 of 5 linked papers with an AI index.
5 papers
ContactFlow: A video action conditioning that transfers across embodiments
Sami Azirar, Enrico Pallotta, Jan Nogga +3
The paper introduces Contact Flow, an embodiment‑agnostic representation that encodes manipulation as the trajectory of 3D contact points, enabling a video‑based world model traine…
Leveraging Vision-Language Models for Open-Vocabulary Instance Segmentation and Tracking
Bastian Pätzold, Jan Nogga, Sven Behnke
Vision-language models (VLMs) excel in visual understanding but often lack reliable grounding capabilities and actionable inference rates. Integrating them with open-vocabulary obj…
Anticipating Human Behavior for Safe Navigation and Efficient Collaborative Manipulation with Mobile Service Robots
Simon Bultmann, Raphael Memmesheimer, Jan Nogga +2
The anticipation of human behavior is a crucial capability for robots to interact with humans safely and efficiently. We employ a smart edge sensor network to provide global observ…
VideoPCDNet: Video Parsing and Prediction with Phase Correlation Networks
Noel José Rodrigues Vicente, Enrique Lehner, Angel Villar-Corrales +2
Understanding and predicting video content is essential for planning and reasoning in dynamic environments. Despite advancements, unsupervised learning of object representations an…
RoboCup@Home 2024 OPL Winner NimbRo: Anthropomorphic Service Robots using Foundation Models for Perception and Planning
Raphael Memmesheimer, Jan Nogga, Bastian Pätzold +8
We present the approaches and contributions of the winning team NimbRo@Home at the RoboCup@Home 2024 competition in the Open Platform League held in Eindhoven, NL. Further, we desc…