12 papers
xperception -- Making Robotic Grasping Easier
Matteo Bortolon, Andrea Caraffa, Alice Fasoli +1
The transition toward high-mix low-volume manufacturing demands flexibility in robotic manipulation. However, conventional vision systems remain a bottleneck, requiring extensive d…
Generative 6D Pose Estimation via Conditional Flow Matching
Amir Hamza, Davide Boscaini, Weihang Li +2
Existing methods for instance-level 6D pose estimation typically rely on neural networks that either directly regress the pose in or estimate it indirectly via loc…
Distilling 3D distinctive local descriptors for 6D pose estimation
Amir Hamza, Andrea Caraffa, Davide Boscaini +1
Three-dimensional local descriptors are crucial for encoding geometric surface properties, making them essential for various point cloud understanding tasks. Among these descriptor…
Leveraging Confident Image Regions for Source-Free Domain-Adaptive Object Detection
Mohamed Lamine Mekhalfi, Davide Boscaini, Fabio Poiesi
Source-free domain-adaptive object detection is an interesting but scarcely addressed topic. It aims at adapting a source-pretrained detector to a distinct target domain without re…
AI-driven visual monitoring of industrial assembly tasks
Mattia Nardon, Stefano Messelodi, Antonio Granata +3
Visual monitoring of industrial assembly tasks is critical for preventing equipment damage due to procedural errors and ensuring worker safety. Although commercial solutions exist,…
An analysis of vision-language models for fabric retrieval
Francesco Giuliari, Asif Khan Pattan, Mohamed Lamine Mekhalfi +1
Effective cross-modal retrieval is essential for applications like information retrieval and recommendation systems, particularly in specialized domains such as manufacturing, wher…