4 citations · 13 across the 8 of their papers we have counts for
8 papers · 1 filter
DTCLMapper: Dual Temporal Consistent Learning for Vectorized HD Map Construction
Siyu Li, Jiacheng Lin, Hao Shi +5
Temporal information plays a pivotal role in Bird's-Eye-View (BEV) driving scene understanding, which can alleviate the visual information sparsity. However, the indiscriminate tem…
MambaMOS: LiDAR-based 3D Moving Object Segmentation with Motion-aware State Space Model
Kang Zeng, Hao Shi, Jiacheng Lin +5
LiDAR-based Moving Object Segmentation (MOS) aims to locate and segment moving objects in point clouds of the current scan using motion information from previous scans. Despite the…
EchoTrack: Auditory Referring Multi-Object Tracking for Autonomous Driving
Jiacheng Lin, Jiajun Chen, Kunyu Peng +4
This paper introduces the task of Auditory Referring Multi-Object Tracking (AR-MOT), which dynamically tracks specific objects in a video sequence based on audio expressions and ap…
Expression Prompt Collaboration Transformer for Universal Referring Video Object Segmentation
Jiajun Chen, Jiacheng Lin, Guojin Zhong +4
Audio-guided Video Object Segmentation (A-VOS) and Referring Video Object Segmentation (R-VOS) are two highly related tasks that both aim to segment specific objects from video seq…
PVPUFormer: Probabilistic Visual Prompt Unified Transformer for Interactive Image Segmentation
Xu Zhang, Kailun Yang, Jiacheng Lin +3
Integration of diverse visual prompts like clicks, scribbles, and boxes in interactive image segmentation significantly facilitates users' interaction as well as improves interacti…
SSD-MonoDETR: Supervised Scale-aware Deformable Transformer for Monocular 3D Object Detection
Xuan He, Fan Yang, Kailun Yang +5
Transformer-based methods have demonstrated superior performance for monocular 3D object detection recently, which aims at predicting 3D attributes from a single 2D image. Most exi…