14 citations · 14 across the 1 of their papers we have counts for
6 papers
Dense Relational Image Captioning via Multi-task Triple-Stream Networks
Dong-Jin Kim, Tae-Hyun Oh, Jinsoo Choi +1
We introduce dense relational captioning, a novel image captioning task which aims to generate multiple captions with respect to relational information between objects in a visual…
Detecting Human-Object Interactions with Action Co-occurrence Priors
Dong-Jin Kim, Xiao Sun, Jinsoo Choi +2
A common problem in human-object interaction (HOI) detection task is that numerous HOI classes have only a small number of labeled examples, resulting in training sets with a long-…
Deep Iterative Frame Interpolation for Full-frame Video Stabilization
Jinsoo Choi, In So Kweon
Video stabilization is a fundamental and important technique for higher quality videos. Prior works have extensively explored video stabilization, but most of them involve cropping…
Image Captioning with Very Scarce Supervised Data: Adversarial Semi-Supervised Learning Approach
Dong-Jin Kim, Jinsoo Choi, Tae-Hyun Oh +1
Constructing an organized dataset comprised of a large number of images and several captions for each image is a laborious task, which requires vast human effort. On the other hand…
Dense Relational Captioning: Triple-Stream Networks for Relationship-Based Captioning
Dong-Jin Kim, Jinsoo Choi, Tae-Hyun Oh +1
Our goal in this work is to train an image captioning model that generates more dense and informative captions. We introduce "relational captioning," a novel image captioning task…
Disjoint Multi-task Learning between Heterogeneous Human-centric Tasks
Dong-Jin Kim, Jinsoo Choi, Tae-Hyun Oh +2
Human behavior understanding is arguably one of the most important mid-level components in artificial intelligence. In order to efficiently make use of data, multi-task learning ha…