30 citations · 37 across the 4 of their papers we have counts for
6 papers
Joint-Motion Mutual Learning for Pose Estimation in Videos
Sifan Wu, Haipeng Chen, Yifang Yin +5
Human pose estimation in videos has long been a compelling yet challenging task within the realm of computer vision. Nevertheless, this task remains difficult because of the comple…
PetalView: Fine-grained Location and Orientation Extraction of Street-view Images via Cross-view Local Search with Supplementary Materials
Wenmiao Hu, Yichen Zhang, Yuxuan Liang +5
Satellite-based street-view information extraction by cross-view matching refers to a task that extracts the location and orientation information of a given street-view image query…
Aircraft Landing Time Prediction with Deep Learning on Trajectory Images
Liping Huang, Sheng Zhang, Yicheng Zhang +2
Aircraft landing time (ALT) prediction is crucial for air traffic management, especially for arrival aircraft sequencing on the runway. In this study, a trajectory image-based deep…
Prototypical Cross-domain Knowledge Transfer for Cervical Dysplasia Visual Inspection
Yichen Zhang, Yifang Yin, Ying Zhang +3
Early detection of dysplasia of the cervix is critical for cervical cancer treatment. However, automatic cervical dysplasia diagnosis via visual inspection, which is more appropria…
Beyond Geo-localization: Fine-grained Orientation of Street-view Images by Cross-view Matching with Satellite Imagery with Supplementary Materials
Wenmiao Hu, Yichen Zhang, Yuxuan Liang +6
Street-view imagery provides us with novel experiences to explore different places remotely. Carefully calibrated street-view images (e.g. Google Street View) can be used for diffe…
Emotionally Enhanced Talking Face Generation
Sahil Goyal, Shagun Uppal, Sarthak Bhagat +3
Several works have developed end-to-end pipelines for generating lip-synced talking faces with various real-world applications, such as teaching and language translation in videos.…