15 citations · 26 across the 14 of their papers we have counts for
23 papers · 1 filter
DiTailed: Ensuring Visual Object Consistency in Text-Image-to-Image Flow Matching Models
Francesco Taioli, Daniel Coelho, Iaroslav Melekhov +4
Despite remarkable progress in text-guided image editing, generative models frequently fail to preserve visual object consistency, defined as the preservation of a subject's key at…
NVSMask3D: Hard Visual Prompting with Camera Pose Interpolation for 3D Open Vocabulary Instance Segmentation
Junyuan Fang, Zihan Wang, Yejun Zhang +3
Vision-language models (VLMs) have demonstrated impressive zero-shot transfer capabilities in image-level visual perception tasks. However, they fall short in 3D instance-level seg…
Road Grip Uncertainty Estimation Through Surface State Segmentation
Jyri Maanpää, Julius Pesonen, Iaroslav Melekhov +2
Slippery road conditions pose significant challenges for autonomous driving. Beyond predicting road grip, it is crucial to estimate its uncertainty reliably to ensure safe vehicle…
A Dataset for Semantic Segmentation in the Presence of Unknowns
Zakaria Laskar, Tomas Vojir, Matej Grcic +6
Before deployment in the real-world deep neural networks require thorough evaluation of how they handle both knowns, inputs represented in the training data, and unknowns (anomalie…
AGS-Mesh: Adaptive Gaussian Splatting and Meshing with Geometric Priors for Indoor Room Reconstruction Using Smartphones
Xuqian Ren, Matias Turkulainen, Jiepeng Wang +4
Geometric priors are often used to enhance 3D reconstruction. With many smartphones featuring low-resolution depth sensors and the prevalence of off-the-shelf monocular geometry es…
Medical Image Segmentation with SAM-generated Annotations
Iira Häkkinen, Iaroslav Melekhov, Erik Englesson +2
The field of medical image segmentation is hindered by the scarcity of large, publicly available annotated datasets. Not all datasets are made public for privacy reasons, and creat…