3 citations · 3 across the 3 of their papers we have counts for
7 papers · 1 filter
SneakPeek: Future-Guided Instructional Streaming Video Generation
Cheeun Hong, German Barquero, Fadime Sener +6
Instructional video generation is an emerging task that aims to synthesize coherent demonstrations of procedural activities from textual descriptions. Such capability has broad imp…
CamCtrl3D: Single-Image Scene Exploration with Precise 3D Camera Control
Stefan Popov, Amit Raj, Michael Krainin +3
We propose a method for generating fly-through videos of a scene, from a single image and a given camera trajectory. We build upon an image-to-video latent diffusion model. We cond…
Efficient Full Image Interactive Segmentation by Leveraging Within-image Appearance Similarity
Mykhaylo Andriluka, Stefano Pellegrini, Stefan Popov +1
We propose a new approach to interactive full-image semantic segmentation which enables quickly collecting training data for new datasets with previously unseen semantic classes (A…
CoReNet: Coherent 3D scene reconstruction from a single RGB image
Stefan Popov, Pablo Bauszat, Vittorio Ferrari
Advances in deep learning techniques have allowed recent work to reconstruct the shape of a single object given only one RBG image as input. Building on common encoder-decoder arch…
C-Flow: Conditional Generative Flow Models for Images and 3D Point Clouds
Albert Pumarola, Stefan Popov, Francesc Moreno-Noguer +1
Flow-based generative models have highly desirable properties like exact log-likelihood evaluation and exact latent-variable inference, however they are still in their infancy and…
Large-scale interactive object segmentation with human annotators
Rodrigo Benenson, Stefan Popov, Vittorio Ferrari
Manually annotating object segmentation masks is very time consuming. Interactive object segmentation methods offer a more efficient alternative where a human annotator and a machi…