30 citations · 146 across the 16 of their papers we have counts for
13 papers · 1 filter
PetalView: Fine-grained Location and Orientation Extraction of Street-view Images via Cross-view Local Search with Supplementary Materials
Wenmiao Hu, Yichen Zhang, Yuxuan Liang +5
Satellite-based street-view information extraction by cross-view matching refers to a task that extracts the location and orientation information of a given street-view image query…
Prototypical Cross-domain Knowledge Transfer for Cervical Dysplasia Visual Inspection
Yichen Zhang, Yifang Yin, Ying Zhang +3
Early detection of dysplasia of the cervix is critical for cervical cancer treatment. However, automatic cervical dysplasia diagnosis via visual inspection, which is more appropria…
SOGDet: Semantic-Occupancy Guided Multi-view 3D Object Detection
Qiu Zhou, Jinming Cao, Hanchao Leng +3
In the field of autonomous driving, accurate and comprehensive perception of the 3D environment is crucial. Bird's Eye View (BEV) based methods have emerged as a promising solution…
In Defense of Clip-based Video Relation Detection
Meng Wei, Long Chen, Wei Ji +2
Video Visual Relation Detection (VidVRD) aims to detect visual relationship triplets in videos using spatial bounding boxes and temporal boundaries. Existing VidVRD methods can be…
Beyond Geo-localization: Fine-grained Orientation of Street-view Images by Cross-view Matching with Satellite Imagery with Supplementary Materials
Wenmiao Hu, Yichen Zhang, Yuxuan Liang +6
Street-view imagery provides us with novel experiences to explore different places remotely. Carefully calibrated street-view images (e.g. Google Street View) can be used for diffe…
Panoptic Scene Graph Generation with Semantics-Prototype Learning
Li Li, Wei Ji, Yiming Wu +4
Panoptic Scene Graph Generation (PSG) parses objects and predicts their relationships (predicate) to connect human language and visual scenes. However, different language preferenc…