49 citations · 72 across the 5 of their papers we have counts for
6 papers
Cross-modal Semantic Enhanced Interaction for Image-Sentence Retrieval
Xuri Ge, Fuhai Chen, Songpei Xu +2
Image-sentence retrieval has attracted extensive research attention in multimedia and computer vision due to its promising application. The key issue lies in jointly learning the v…
Factored Attention and Embedding for Unstructured-view Topic-related Ultrasound Report Generation
Fuhai Chen, Rongrong Ji, Chengpeng Dai +4
Echocardiography is widely used to clinical practice for diagnosis and treatment, e.g., on the common congenital heart defects. The traditional manual manipulation is error-prone d…
Structured Multi-modal Feature Embedding and Alignment for Image-Sentence Retrieval
Xuri Ge, Fuhai Chen, Joemon M. Jose +3
The current state-of-the-art image-sentence retrieval methods implicitly align the visual-textual fragments, like regions in images and words in sentences, and adopt attention modu…
Improving Image Captioning by Leveraging Intra- and Inter-layer Global Representation in Transformer Network
Jiayi Ji, Yunpeng Luo, Xiaoshuai Sun +5
Transformer-based architectures have shown great success in image captioning, where object regions are encoded and then attended into the vectorial representations to guide the cap…
Semantic-aware Image Deblurring
Fuhai Chen, Rongrong Ji, Chengpeng Dai +6
Image deblurring has achieved exciting progress in recent years. However, traditional methods fail to deblur severely blurred images, where semantic contents appears ambiguously. I…
Scene-based Factored Attention for Image Captioning
Chen Shen, Rongrong Ji, Fuhai Chen +2
Image captioning has attracted ever-increasing research attention in the multimedia community. To this end, most cutting-edge works rely on an encoder-decoder framework with attent…