23 citations · 74 across the 13 of their papers we have counts for
5 papers · 2 filters
Consensus-based Sequence Training for Video Captioning
Sang Phan, Gustav Eje Henter, Yusuke Miyao +1
Captioning models are typically trained using the cross-entropy loss. However, their performance is evaluated on other metrics designed to better correlate with human assessments.…
Discriminative Learning of Open-Vocabulary Object Retrieval and Localization by Negative Phrase Augmentation
Ryota Hinami, Shin'ichi Satoh
Thanks to the success of object detection technology, we can retrieve objects of the specified classes even from huge image collections. However, the current state-of-the-art objec…
Joint Detection and Recounting of Abnormal Events by Learning Deep Generic Knowledge
Ryota Hinami, Tao Mei, Shin'ichi Satoh
This paper addresses the problem of joint detection and recounting of abnormal events in videos. Recounting of abnormal events, i.e., explaining why they are judged to be abnormal,…
Active Learning for Structured Prediction from Partially Labeled Data
Mehran Khodabandeh, Zhiwei Deng, Mostafa S. Ibrahim +2
We propose a general purpose active learning algorithm for structured prediction, gathering labeled data for training a model that outputs a set of related labels for an image or v…
Embedding Watermarks into Deep Neural Networks
Yusuke Uchida, Yuki Nagai, Shigeyuki Sakazawa +1
Deep neural networks have recently achieved significant progress. Sharing trained models of these deep neural networks is very important in the rapid progress of researching or dev…