1 citations · 1 across the 4 of their papers we have counts for
6 papers · 1 filter
A Simple Recipe for Language-guided Domain Generalized Segmentation
Mohammad Fahes, Tuan-Hung Vu, Andrei Bursuc +2
Generalization to new domains not seen during training is one of the long-standing challenges in deploying neural networks in real-world applications. Existing generalization techn…
Decaf: Monocular Deformation Capture for Face and Hand Interactions
Soshi Shimada, Vladislav Golyanik, Patrick Pérez +1
Existing methods for 3D tracking from monocular RGB videos predominantly consider articulated and rigid objects. Modelling dense non-rigid object deformations in this setting remai…
Three Pillars improving Vision Foundation Model Distillation for Lidar
Gilles Puy, Spyros Gidaris, Alexandre Boulch +5
Self-supervised image backbones can be used to address complex 2D tasks (e.g., semantic segmentation, object discovery) very efficiently and with little or no downstream supervisio…
Unsupervised Object Localization in the Era of Self-Supervised ViTs: A Survey
Oriane Siméoni, Éloi Zablocki, Spyros Gidaris +2
The recent enthusiasm for open-world vision systems show the high interest of the community to perform perception tasks outside of the closed-vocabulary benchmark setups which have…
T-UDA: Temporal Unsupervised Domain Adaptation in Sequential Point Clouds
Awet Haileslassie Gebrehiwot, David Hurych, Karel Zimmermann +2
Deep perception models have to reliably cope with an open-world setting of domain shifts induced by different geographic regions, sensor properties, mounting positions, and several…
DiffHPE: Robust, Coherent 3D Human Pose Lifting with Diffusion
Cédric Rommel, Eduardo Valle, Mickaël Chen +4
We present an innovative approach to 3D Human Pose Estimation (3D-HPE) by integrating cutting-edge diffusion models, which have revolutionized diverse fields, but are relatively un…