most citedExpression Prompt Collaboration Transformer for Universal Referring Video Object Segmentation

4 citations · 13 across the 8 of their papers we have counts for

collaborators
Showing cs.CVShow all

8 papers · 1 filter

cs.CV2024

DTCLMapper: Dual Temporal Consistent Learning for Vectorized HD Map Construction

Siyu Li, Jiacheng Lin, Hao Shi +5

Temporal information plays a pivotal role in Bird's-Eye-View (BEV) driving scene understanding, which can alleviate the visual information sparsity. However, the indiscriminate tem…

cs.CV2024★ 1 cited

MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model

Kang Zeng, Hao Shi, Jiacheng Lin +5

LiDAR-based Moving Object Segmentation (MOS) aims to locate and segment moving objects in point clouds of the current scan using motion information from previous scans. Despite the…

cs.CV2024

EchoTrack: Auditory Referring Multi-Object Tracking for Autonomous Driving

Jiacheng Lin, Jiajun Chen, Kunyu Peng +4

This paper introduces the task of Auditory Referring Multi-Object Tracking (AR-MOT), which dynamically tracks specific objects in a video sequence based on audio expressions and ap…

cs.CV2023★ 4 cited

Expression Prompt Collaboration Transformer for Universal Referring Video Object Segmentation

Jiajun Chen, Jiacheng Lin, Guojin Zhong +4

Audio-guided Video Object Segmentation (A-VOS) and Referring Video Object Segmentation (R-VOS) are two highly related tasks that both aim to segment specific objects from video seq…

cs.CV2023★ 4 cited

PVPUFormer: Probabilistic Visual Prompt Unified Transformer for Interactive Image Segmentation

Xu Zhang, Kailun Yang, Jiacheng Lin +3

Integration of diverse visual prompts like clicks, scribbles, and boxes in interactive image segmentation significantly facilitates users' interaction as well as improves interacti…

cs.CV2023

SSD-MonoDETR: Supervised Scale-aware Deformable Transformer for Monocular 3D Object Detection

Xuan He, Fan Yang, Kailun Yang +5

Transformer-based methods have demonstrated superior performance for monocular 3D object detection recently, which aims at predicting 3D attributes from a single 2D image. Most exi…