2 citations · 3 across the 5 of their papers we have counts for
8 papers · 1 filter
A Hierarchical Semantic Distillation Framework for Open-Vocabulary Object Detection
Shenghao Fu, Junkai Yan, Qize Yang +3
Open-vocabulary object detection (OVD) aims to detect objects beyond the training annotations, where detectors are usually aligned to a pre-trained vision-language model, eg, CLIP,…
LLMDet: Learning Strong Open-Vocabulary Object Detectors under the Supervision of Large Language Models
Shenghao Fu, Qize Yang, Qijie Mo +5
Recent open-vocabulary detectors achieve promising performance with abundant region-level annotated data. In this work, we show that an open-vocabulary detector co-training with a…
Frozen-DETR: Enhancing DETR with Image Understanding from Frozen Foundation Models
Shenghao Fu, Junkai Yan, Qize Yang +3
Recent vision foundation models can extract universal representations and show impressive abilities in various tasks. However, their application on object detection is largely over…
Loc4Plan: Locating Before Planning for Outdoor Vision and Language Navigation
Huilin Tian, Jingke Meng, Wei-Shi Zheng +3
Vision and Language Navigation (VLN) is a challenging task that requires agents to understand instructions and navigate to the destination in a visual environment.One of the key ch…
Bridge Past and Future: Overcoming Information Asymmetry in Incremental Object Detection
Qijie Mo, Yipeng Gao, Shenghao Fu +3
In incremental object detection, knowledge distillation has been proven to be an effective way to alleviate catastrophic forgetting. However, previous works focused on preserving t…
DreamView: Injecting View-specific Text Guidance into Text-to-3D Generation
Junkai Yan, Yipeng Gao, Qize Yang +4
Text-to-3D generation, which synthesizes 3D assets according to an overall text description, has significantly progressed. However, a challenge arises when the specific appearances…