2 citations · 2 across the 3 of their papers we have counts for
3 papers
cs.CV2024
Text-guided Zero-Shot Object Localization
Jingjing Wang, Xinglin Piao, Zongzhi Gao +3
Object localization is a hot issue in computer vision area, which aims to identify and determine the precise location of specific objects from image or video. Most existing object…
cs.CV2024★ 2 cited
UIFormer: A Unified Transformer-based Framework for Incremental Few-Shot Object Detection and Instance Segmentation
Chengyuan Zhang, Yilin Zhang, Lei Zhu +6
This paper introduces a novel framework for unified incremental few-shot object detection (iFSOD) and instance segmentation (iFSIS) using the Transformer architecture. Our goal is…
cs.CV2024
Mono-ViFI: A Unified Learning Framework for Self-supervised Single- and Multi-frame Monocular Depth Estimation
Jinfeng Liu, Lingtong Kong, Bo Li +3
Self-supervised monocular depth estimation has gathered notable interest since it can liberate training from dependency on depth annotations. In monocular video training case, rece…