22 citations · 53 across the 8 of their papers we have counts for
8 papers
Cross-Modality Domain Adaptation for Freespace Detection: A Simple yet Effective Baseline
Yuanbin Wang, Leyan Zhu, Shaofei Huang +4
As one of the fundamental functions of autonomous driving system, freespace detection aims at classifying each pixel of the image captured by the camera as drivable or non-drivable…
A Keypoint-based Global Association Network for Lane Detection
Jinsheng Wang, Yinchao Ma, Shaofei Huang +4
Lane detection is a challenging task that requires predicting complex topology shapes of lane lines and distinguishing different types of lanes simultaneously. Earlier works follow…
TransRefer3D: Entity-and-Relation Aware Transformer for Fine-Grained 3D Visual Grounding
Dailan He, Yusheng Zhao, Junyu Luo +4
Recently proposed fine-grained 3D visual grounding is an essential and challenging task, whose goal is to identify the 3D object referred by a natural language sentence from other…
Cross-Modal Progressive Comprehension for Referring Segmentation
Si Liu, Tianrui Hui, Shaofei Huang +3
Given a natural language expression and an image/video, the goal of referring segmentation is to produce the pixel-level masks of the entities described by the subject of the expre…
Collaborative Spatial-Temporal Modeling for Language-Queried Video Actor Segmentation
Tianrui Hui, Shaofei Huang, Si Liu +5
Language-queried video actor segmentation aims to predict the pixel-level mask of the actor which performs the actions described by a natural language query in the target frames. E…
ORDNet: Capturing Omni-Range Dependencies for Scene Parsing
Shaofei Huang, Si Liu, Tianrui Hui +4
Learning to capture dependencies between spatial positions is essential to many visual tasks, especially the dense labeling problems like scene parsing. Existing methods can effect…