4 papers · 1 filter
Open-Vocabulary BEV Segmentation with 3D-Aware Geometric Constraints
Hojun Choi, Seulbin Hwang, Daejung Kim +4
Bird's-eye view (BEV) perception fuses multi-camera images into a unified top-down representation for autonomous driving. Despite recent progress, state-of-the-art methods remain c…
MSPL: Multi-Step Pseudo-Labeling for Open-Vocabulary Object Detection
Hojun Choi, Youngsun Lim, Jaeyo Shin +1
Open-vocabulary object detection (OVD) aims to recognize and localize object categories beyond the training set. Recent approaches leverage vision-language models to generate pseud…
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering
Youngsun Lim, Hojun Choi, Hyunjung Shim
Despite the impressive success of text-to-image (TTI) generation models, existing studies overlook the issue of whether these models accurately convey factual information. In this…
Sampling Bag of Views for Open-Vocabulary Object Detection
Hojun Choi, Junsuk Choe, Hyunjung Shim
Existing open-vocabulary object detection (OVD) develops methods for testing unseen categories by aligning object region embeddings with corresponding VLM features. A recent study…