activity
20242026
most citedLeveraging Vision-Language Models for Open-Vocabulary Instance Segmentation and Tracking

2 citations · 3 across the 16 of their papers we have counts for

collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV2026

Efficient Image Annotation via Semi-Supervised Object Segmentation with Label Propagation

Vitalii Tutevych, Raphael Memmesheimer, Luca Eichler +4

Reliable object perception is necessary for general-purpose service robots. Open-vocabulary detectors struggle to generalize beyond a few classes and fully supervised training of o…

cs.CV2025★ 2 cited

Leveraging Vision-Language Models for Open-Vocabulary Instance Segmentation and Tracking

Bastian Pätzold, Jan Nogga, Sven Behnke

Vision-language models (VLMs) excel in visual understanding but often lack reliable grounding capabilities and actionable inference rates. Integrating them with open-vocabulary obj…

cs.CV2025

LIAM: Multimodal Transformer for Language Instructions, Images, Actions and Semantic Maps

Yihao Wang, Raphael Memmesheimer, Sven Behnke

The availability of large language models and open-vocabulary object perception methods enables more flexibility for domestic service robots. The large variability of domestic task…

cs.CV2024

Person Segmentation and Action Classification for Multi-Channel Hemisphere Field of View LiDAR Sensors

Svetlana Seliunina, Artem Otelepko, Raphael Memmesheimer +1

Robots need to perceive persons in their surroundings for safety and to interact with them. In this paper, we present a person segmentation and action classification approach that…

cs.CV2024

Marker-free Human Gait Analysis using a Smart Edge Sensor System

Eva Katharina Bauer, Simon Bultmann, Sven Behnke

The human gait is a complex interplay between the neuronal and the muscular systems, reflecting an individual's neurological and physiological condition. This makes gait analysis a…