48 citations · 80 across the 5 of their papers we have counts for
6 papers · 1 filter
Controllable Inversion of Black-Box Face Recognition Models via Diffusion
Manuel Kansy, Anton Raël, Graziana Mignone +4
Face recognition models embed a face image into a low-dimensional identity vector containing abstract encodings of identity-specific facial features that allow individuals to be di…
Lossy Image Compression with Normalizing Flows
Leonhard Helminger, Abdelaziz Djelouah, Markus Gross +1
Deep learning based image compression has recently witnessed exciting progress and in some cases even managed to surpass transform coding based approaches that have been establishe…
Enriching Video Captions With Contextual Text
Philipp Rimle, Pelin Dogan, Markus Gross
Understanding video content and generating caption with context is an important and challenging task. Unlike prior methods that typically attempt to generate generic video captions…
Neural Sequential Phrase Grounding (SeqGROUND)
Pelin Dogan, Leonid Sigal, Markus Gross
We propose an end-to-end approach for phrase grounding in images. Unlike prior methods that typically attempt to ground each phrase independently by building an image-text embeddin…
Deep Video Color Propagation
Simone Meyer, Victor Cornillère, Abdelaziz Djelouah +2
Traditional approaches for color propagation in videos rely on some form of matching between consecutive video frames. Using appearance descriptors, colors are then propagated both…
A Neural Multi-sequence Alignment TeCHnique (NeuMATCH)
Pelin Dogan, Boyang Li, Leonid Sigal +1
The alignment of heterogeneous sequential data (video to text) is an important and challenging problem. Standard techniques for this task, including Dynamic Time Warping (DTW) and…