1 citations · 1 across the 3 of their papers we have counts for
3 papers · 1 filter
Zero-Shot Character Identification and Speaker Prediction in Comics via Iterative Multimodal Fusion
Yingxuan Li, Ryota Hinami, Kiyoharu Aizawa +1
Recognizing characters and predicting speakers of dialogue are critical for comic processing tasks, such as voice generation or translation. However, because characters vary by com…
Multimodal Co-Training for Selecting Good Examples from Webly Labeled Video
Ryota Hinami, Junwei Liang, Shin'ichi Satoh +1
We tackle the problem of learning concept classifiers from videos on the web without using manually labeled data. Although metadata attached to videos (e.g., video titles, descript…
Region-Based Image Retrieval Revisited
Ryota Hinami, Yusuke Matsui, Shin'ichi Satoh
Region-based image retrieval (RBIR) technique is revisited. In early attempts at RBIR in the late 90s, researchers found many ways to specify region-based queries and spatial relat…