23 citations · 69 across the 10 of their papers we have counts for
4 papers · 1 filter
Consensus-based Sequence Training for Video Captioning
Sang Phan, Gustav Eje Henter, Yusuke Miyao +1
Captioning models are typically trained using the cross-entropy loss. However, their performance is evaluated on other metrics designed to better correlate with human assessments.…
Joint Detection and Recounting of Abnormal Events by Learning Deep Generic Knowledge
Ryota Hinami, Tao Mei, Shin'ichi Satoh
This paper addresses the problem of joint detection and recounting of abnormal events in videos. Recounting of abnormal events, i.e., explaining why they are judged to be abnormal,…
Region-Based Image Retrieval Revisited
Ryota Hinami, Yusuke Matsui, Shin'ichi Satoh
Region-based image retrieval (RBIR) technique is revisited. In early attempts at RBIR in the late 90s, researchers found many ways to specify region-based queries and spatial relat…
Active Learning for Structured Prediction from Partially Labeled Data
Mehran Khodabandeh, Zhiwei Deng, Mostafa S. Ibrahim +2
We propose a general purpose active learning algorithm for structured prediction, gathering labeled data for training a model that outputs a set of related labels for an image or v…