activity
20232025
most citedLLMDet: Learning Strong Open-Vocabulary Object Detectors under the Supervision of Large Language Models

2 citations · 3 across the 5 of their papers we have counts for

collaborators
Showing cs.CVShow all

8 papers · 1 filter

cs.CV2025

A Hierarchical Semantic Distillation Framework for Open-Vocabulary Object Detection

Shenghao Fu, Junkai Yan, Qize Yang +3

Open-vocabulary object detection (OVD) aims to detect objects beyond the training annotations, where detectors are usually aligned to a pre-trained vision-language model, eg, CLIP,…

cs.CV20252 cited

LLMDet: Learning Strong Open-Vocabulary Object Detectors under the Supervision of Large Language Models

Shenghao Fu, Qize Yang, Qijie Mo +5

Recent open-vocabulary detectors achieve promising performance with abundant region-level annotated data. In this work, we show that an open-vocabulary detector co-training with a…

cs.CV2024

Frozen-DETR: Enhancing DETR with Image Understanding from Frozen Foundation Models

Shenghao Fu, Junkai Yan, Qize Yang +3

Recent vision foundation models can extract universal representations and show impressive abilities in various tasks. However, their application on object detection is largely over…

cs.CV2024

Loc4Plan: Locating Before Planning for Outdoor Vision and Language Navigation

Huilin Tian, Jingke Meng, Wei-Shi Zheng +3

Vision and Language Navigation (VLN) is a challenging task that requires agents to understand instructions and navigate to the destination in a visual environment.One of the key ch…

cs.CV2024

Bridge Past and Future: Overcoming Information Asymmetry in Incremental Object Detection

Qijie Mo, Yipeng Gao, Shenghao Fu +3

In incremental object detection, knowledge distillation has been proven to be an effective way to alleviate catastrophic forgetting. However, previous works focused on preserving t…

cs.CV2024

DreamView: Injecting View-specific Text Guidance into Text-to-3D Generation

Junkai Yan, Yipeng Gao, Qize Yang +4

Text-to-3D generation, which synthesizes 3D assets according to an overall text description, has significantly progressed. However, a challenge arises when the specific appearances…