17 citations · 19 across the 2 of their papers we have counts for
4 papers
Voice2Mesh: Cross-Modal 3D Face Model Generation from Voices
Cho-Ying Wu, Ke Xu, Chin-Cheng Hsu +1
This work focuses on the analysis that whether 3D face models can be learned from only the speech inputs of speakers. Previous works for cross-modal face synthesis study image gene…
DeepMask: an algorithm for cloud and cloud shadow detection in optical satellite remote sensing images using deep residual network
Ke Xu, Kaiyu Guan, Jian Peng +2
Detecting and masking cloud and cloud shadow from satellite remote sensing images is a pervasive problem in the remote sensing community. Accurate and efficient detection of cloud…
Online monitoring for safe pedestrian-vehicle interactions
Peter Du, Zhe Huang, Tianqi Liu +5
As autonomous systems begin to operate amongst humans, methods for safe interaction must be investigated. We consider an example of a small autonomous vehicle in a pedestrian zone…
Revisiting Image-Language Networks for Open-ended Phrase Detection
Bryan A. Plummer, Kevin J. Shih, Yichen Li +4
Most existing work that grounds natural language phrases in images starts with the assumption that the phrase in question is relevant to the image. In this paper we address a more…