activity
20152023
most citedCityFlow-NL: Tracking and Retrieval of Vehicles at City Scale by Natural Language Descriptions

27 citations · 113 across the 21 of their papers we have counts for

collaborators
Showing 2020 · cs.CVShow all

11 papers · 2 filters

cs.CV2020

Self-supervised Visual Attribute Learning for Fashion Compatibility

Donghyun Kim, Kuniaki Saito, Samarth Mishra +3

Many self-supervised learning (SSL) methods have been successful in learning semantically meaningful visual representations by solving pretext tasks. However, prior work in SSL foc…

cs.CV2020★ 6 cited

Real-time Semantic Segmentation with Fast Attention

Ping Hu, Federico Perazzi, Fabian Caba Heilbron +4

In deep CNN based models for semantic segmentation, high accuracy relies on rich spatial context (large receptive fields) and fine spatial details (high resolution), both of which…

cs.CV2020★ 8 cited

Protecting Against Image Translation Deepfakes by Leaking Universal Perturbations from Black-Box Neural Networks

Nataniel Ruiz, Sarah Adel Bargal, Stan Sclaroff

In this work, we develop efficient disruptions of black-box image translation deepfake generation systems. We are the first to demonstrate black-box deepfake generation disruption…

cs.CV2020★ 1 cited

Disrupting Deepfakes: Adversarial Attacks Against Conditional Image Translation Networks and Facial Manipulation Systems

Nataniel Ruiz, Sarah Adel Bargal, Stan Sclaroff

Face modification systems using deep learning have become increasingly powerful and accessible. Given images of a person's face, such systems can generate new images of that same p…

cs.CV2020★ 10 cited

Temporally Distributed Networks for Fast Video Semantic Segmentation

Ping Hu, Fabian Caba Heilbron, Oliver Wang +3

We present TDNet, a temporally distributed network designed for fast and accurate video semantic segmentation. We observe that features extracted from a certain high-level layer of…

cs.CV2020

Spatio-Temporal Action Detection with Multi-Object Interaction

Huijuan Xu, Lizhi Yang, Stan Sclaroff +2

Spatio-temporal action detection in videos requires localizing the action both spatially and temporally in the form of an "action tube". Nowadays, most spatio-temporal action detec…