7k citations
- University of California, Santa BarbaraUS109 papers
- Microsoft Research (United Kingdom)GB45 papers
- University of California, BerkeleyUS45 papers
- ETH ZurichCH44 papers
- University of Maryland, College ParkUS43 papers
- Carnegie Mellon UniversityUS39 papers
- Stanford UniversityUS39 papers
- University of WashingtonUS35 papers
- Cornell UniversityUS32 papers
- Princeton UniversityUS31 papers
- California Institute of TechnologyUS28 papers
- Georgia Institute of TechnologyUS24 papers
54 papers · 1 filter
Video Summarization Overview
Mayu Otani, Yale Song, Yang Wang
With the broad growth of video capturing devices and applications on the web, it is more demanding to provide desired video content for users efficiently. Video summarization facil…
Backdoor Attacks on Crowd Counting
Yuhua Sun, Tailai Zhang, Xingjun Ma +6
Crowd counting is a regression task that estimates the number of people in a scene image, which plays a vital role in a range of safety-critical applications, such as video surveil…
MobilePhys: Personalized Mobile Camera-Based Contactless Physiological Sensing
Xin Liu, Yuntao Wang, Sinan Xie +4
Camera-based contactless photoplethysmography refers to a set of popular techniques for contactless physiological measurement. The current state-of-the-art neural models are typica…
Improving Visual Quality of Image Synthesis by A Token-based Generator with Transformers
Yanhong Zeng, Huan Yang, Hongyang Chao +2
We present a new perspective of achieving image synthesis by viewing this task as a visual token generation problem. Different from existing paradigms that directly synthesize a fu…
Bootstrap Your Object Detector via Mixed Training
Mengde Xu, Zheng Zhang, Fangyun Wei +5
We introduce MixTraining, a new training paradigm for object detection that can improve the performance of existing detectors for free. MixTraining enhances data augmentation by ut…
SOAT: A Scene- and Object-Aware Transformer for Vision-and-Language Navigation
Abhinav Moudgil, Arjun Majumdar, Harsh Agrawal +2
Natural language instructions for visual navigation often use scene descriptions (e.g., "bedroom") and object references (e.g., "green chairs") to provide a breadcrumb trail to a g…