activity
20152023
most citedCascade R-CNN: Delving into High Quality Object Detection

398 citations · 766 across the 26 of their papers we have counts for

collaborators
Showing cs.CVShow all

38 papers · 1 filter

cs.CV20233 cited

POP: Prompt Of Prompts for Continual Learning

Zhiyuan Hu, Jiancheng Lyu, Dashan Gao +1

Continual learning (CL) has attracted increasing attention in the recent past. It aims to mimic the human ability to learn new concepts without catastrophic forgetting. While exist…

cs.CV20221 cited

YORO -- Lightweight End to End Visual Grounding

Chih-Hui Ho, Srikar Appalaraju, Bhavan Jasani +2

We present YORO - a multi-modal transformer encoder-only architecture for the Visual Grounding (VG) task. This task involves localizing, in an image, an object referred via natural…

cs.CV2022

CoordGAN: Self-Supervised Dense Correspondences Emerge from GANs

Jiteng Mu, Shalini De Mello, Zhiding Yu +4

Recent advances show that Generative Adversarial Networks (GANs) can synthesize images with smooth variations along semantically meaningful latent directions, such as pose, express…

cs.CV20222 cited

Omni-DETR: Omni-Supervised Object Detection with Transformers

Pei Wang, Zhaowei Cai, Hao Yang +4

We consider the problem of omni-supervised object detection, which can use unlabeled, fully labeled and weakly labeled annotations, such as image tags, counts, points, etc., for ob…

cs.CV2021

Black-Box Test-Time Shape REFINEment for Single View 3D Reconstruction

Brandon Leung, Chih-Hui Ho, Nuno Vasconcelos

Much recent progress has been made in reconstructing the 3D shape of an object from an image of it, i.e. single view 3D reconstruction. However, it has been suggested that current…

cs.CV2021

Learning of Visual Relations: The Devil is in the Tails

Alakh Desai, Tz-Ying Wu, Subarna Tripathi +1

Significant effort has been recently devoted to modeling visual relations. This has mostly addressed the design of architectures, typically by adding parameters and increasing mode…