49 citations · 55 across the 4 of their papers we have counts for
4 papers
Cross-modal Semantic Enhanced Interaction for Image-Sentence Retrieval
Xuri Ge, Fuhai Chen, Songpei Xu +2
Image-sentence retrieval has attracted extensive research attention in multimedia and computer vision due to its promising application. The key issue lies in jointly learning the v…
Automatic Facial Paralysis Estimation with Facial Action Units
Xuri Ge, Joemon M. Jose, Pengcheng Wang +3
Facial palsy is unilateral facial nerve weakness or paralysis of rapid onset with unknown causes. Automatically estimating facial palsy severeness can be helpful for the diagnosis…
Factored Attention and Embedding for Unstructured-view Topic-related Ultrasound Report Generation
Fuhai Chen, Rongrong Ji, Chengpeng Dai +4
Echocardiography is widely used to clinical practice for diagnosis and treatment, e.g., on the common congenital heart defects. The traditional manual manipulation is error-prone d…
Structured Multi-modal Feature Embedding and Alignment for Image-Sentence Retrieval
Xuri Ge, Fuhai Chen, Joemon M. Jose +3
The current state-of-the-art image-sentence retrieval methods implicitly align the visual-textual fragments, like regions in images and words in sentences, and adopt attention modu…