output
20022024
most citedNon-Abelian Anyons and Topological Quantum Computation

7k citations

Showing cs.CVShow all

54 papers · 1 filter

cs.CV202212 cited

Video Summarization Overview

Mayu Otani, Yale Song, Yang Wang

With the broad growth of video capturing devices and applications on the web, it is more demanding to provide desired video content for users efficiently. Video summarization facil…

cs.CV202211 cited

Backdoor Attacks on Crowd Counting

Yuhua Sun, Tailai Zhang, Xingjun Ma +6

Crowd counting is a regression task that estimates the number of people in a scene image, which plays a vital role in a range of safety-critical applications, such as video surveil…

cs.CV202220 cited

MobilePhys: Personalized Mobile Camera-Based Contactless Physiological Sensing

Xin Liu, Yuntao Wang, Sinan Xie +4

Camera-based contactless photoplethysmography refers to a set of popular techniques for contactless physiological measurement. The current state-of-the-art neural models are typica…

cs.CV20215 cited

Improving Visual Quality of Image Synthesis by A Token-based Generator with Transformers

Yanhong Zeng, Huan Yang, Hongyang Chao +2

We present a new perspective of achieving image synthesis by viewing this task as a visual token generation problem. Different from existing paradigms that directly synthesize a fu…

cs.CV2021

Bootstrap Your Object Detector via Mixed Training

Mengde Xu, Zheng Zhang, Fangyun Wei +5

We introduce MixTraining, a new training paradigm for object detection that can improve the performance of existing detectors for free. MixTraining enhances data augmentation by ut…

cs.CV20214 cited

SOAT: A Scene- and Object-Aware Transformer for Vision-and-Language Navigation

Abhinav Moudgil, Arjun Majumdar, Harsh Agrawal +2

Natural language instructions for visual navigation often use scene descriptions (e.g., "bedroom") and object references (e.g., "green chairs") to provide a breadcrumb trail to a g…