most citedLearning Object-Language Alignments for Open-Vocabulary Object Detection

36 citations · 58 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CV202236 cited

Learning Object-Language Alignments for Open-Vocabulary Object Detection

Chuang Lin, Peize Sun, Yi Jiang +5

Existing object detection methods are bounded in a fixed-set vocabulary by costly labeled data. When dealing with novel categories, the model has to be retrained with more bounding…

cs.CV20221 cited

ExtrudeNet: Unsupervised Inverse Sketch-and-Extrude for Shape Parsing

Daxuan Ren, Jianmin Zheng, Jianfei Cai +2

Sketch-and-extrude is a common and intuitive modeling process in computer aided design. This paper studies the problem of learning the shape given in the form of point clouds by in…

cs.CV202214 cited

MoVQ: Modulating Quantized Vectors for High-Fidelity Image Generation

Chuanxia Zheng, Long Tung Vuong, Jianfei Cai +1

Although two-stage Vector Quantized (VQ) generative models allow for synthesizing high-fidelity and high-resolution images, their quantization operator encodes similar patches with…

cs.CV2022

Transformer Scale Gate for Semantic Segmentation

Hengcan Shi, Munawar Hayat, Jianfei Cai

Effectively encoding multi-scale contextual information is crucial for accurate semantic segmentation. Existing transformer-based segmentation models combine features across scales…

cs.CV20227 cited

High-Quality Pluralistic Image Completion via Code Shared VQGAN

Chuanxia Zheng, Guoxian Song, Tat-Jen Cham +3

PICNet pioneered the generation of multiple and diverse results for image completion task, but it required a careful balance between loss (diversity) and reconstruct…