output
20022026
most citedNon-Abelian Anyons and Topological Quantum Computation

7k citations

Showing 2021 · cs.CVShow all

12 papers · 2 filters

cs.CV20215 cited

Improving Visual Quality of Image Synthesis by A Token-based Generator with Transformers

Yanhong Zeng, Huan Yang, Hongyang Chao +2

We present a new perspective of achieving image synthesis by viewing this task as a visual token generation problem. Different from existing paradigms that directly synthesize a fu…

cs.CV2021

Bootstrap Your Object Detector via Mixed Training

Mengde Xu, Zheng Zhang, Fangyun Wei +5

We introduce MixTraining, a new training paradigm for object detection that can improve the performance of existing detectors for free. MixTraining enhances data augmentation by ut…

cs.CV20214 cited

SOAT: A Scene- and Object-Aware Transformer for Vision-and-Language Navigation

Abhinav Moudgil, Arjun Majumdar, Harsh Agrawal +2

Natural language instructions for visual navigation often use scene descriptions (e.g., "bedroom") and object references (e.g., "green chairs") to provide a breadcrumb trail to a g…

cs.CV202114 cited

Semi-Supervised Semantic Segmentation via Adaptive Equalization Learning

Hanzhe Hu, Fangyun Wei, Han Hu +3

Due to the limited and even imbalanced data, semi-supervised semantic segmentation tends to have poor performance on some certain categories, e.g., tailed categories in Cityscapes…

cs.CV2021

PatchMatch-RL: Deep MVS with Pixelwise Depth, Normal, and Visibility

Jae Yong Lee, Joseph DeGol, Chuhang Zou +1

Recent learning-based multi-view stereo (MVS) methods show excellent performance with dense cameras and small depth ranges. However, non-learning based approaches still outperform…

cs.CV2021

PoseRN: A 2D pose refinement network for bias-free multi-view 3D human pose estimation

Akihiko Sayo, Diego Thomas, Hiroshi Kawasaki +2

We propose a new 2D pose refinement network that learns to predict the human bias in the estimated 2D pose. There are biases in 2D pose estimations that are due to differences betw…