4 citations · 4 across the 1 of their papers we have counts for
1 paper · 1 filter
Yang Xian, Yingli Tian
In this paper, a self-guiding multimodal LSTM (sg-LSTM) image captioning model is proposed to handle uncontrolled imbalanced real-world image-sentence dataset. We collect FlickrNYC…