activity
20182024
most citedBeyond Geo-localization: Fine-grained Orientation of Street-view Images by Cross-view Matching with Satellite Imagery with Supplementary Materials

30 citations · 146 across the 16 of their papers we have counts for

collaborators
Showing cs.CVShow all

13 papers · 1 filter

cs.CV20244 cited

PetalView: Fine-grained Location and Orientation Extraction of Street-view Images via Cross-view Local Search with Supplementary Materials

Wenmiao Hu, Yichen Zhang, Yuxuan Liang +5

Satellite-based street-view information extraction by cross-view matching refers to a task that extracts the location and orientation information of a given street-view image query…

cs.CV20234 cited

Prototypical Cross-domain Knowledge Transfer for Cervical Dysplasia Visual Inspection

Yichen Zhang, Yifang Yin, Ying Zhang +3

Early detection of dysplasia of the cervix is critical for cervical cancer treatment. However, automatic cervical dysplasia diagnosis via visual inspection, which is more appropria…

cs.CV2023

SOGDet: Semantic-Occupancy Guided Multi-view 3D Object Detection

Qiu Zhou, Jinming Cao, Hanchao Leng +3

In the field of autonomous driving, accurate and comprehensive perception of the 3D environment is crucial. Bird's Eye View (BEV) based methods have emerged as a promising solution…

cs.CV2023

In Defense of Clip-based Video Relation Detection

Meng Wei, Long Chen, Wei Ji +2

Video Visual Relation Detection (VidVRD) aims to detect visual relationship triplets in videos using spatial bounding boxes and temporal boundaries. Existing VidVRD methods can be…

cs.CV202330 cited

Beyond Geo-localization: Fine-grained Orientation of Street-view Images by Cross-view Matching with Satellite Imagery with Supplementary Materials

Wenmiao Hu, Yichen Zhang, Yuxuan Liang +6

Street-view imagery provides us with novel experiences to explore different places remotely. Carefully calibrated street-view images (e.g. Google Street View) can be used for diffe…

cs.CV2023

Panoptic Scene Graph Generation with Semantics-Prototype Learning

Li Li, Wei Ji, Yiming Wu +4

Panoptic Scene Graph Generation (PSG) parses objects and predicts their relationships (predicate) to connect human language and visual scenes. However, different language preferenc…