4 papers · 1 filter
Cap2Det: Learning to Amplify Weak Caption Supervision for Object Detection
Keren Ye, Mingda Zhang, Adriana Kovashka +3
Learning to localize and name object instances is a fundamental problem in vision, but state-of-the-art approaches rely on expensive bounding box supervision. While weakly supervis…
Learning to discover and localize visual objects with open vocabulary
Keren Ye, Mingda Zhang, Wei Li +3
To alleviate the cost of obtaining accurate bounding boxes for training today's state-of-the-art object detection models, recent weakly supervised detection work has proposed techn…
Equal But Not The Same: Understanding the Implicit Relationship Between Persuasive Images and Text
Mingda Zhang, Rebecca Hwa, Adriana Kovashka
Images and text in advertisements interact in complex, non-literal ways. The two channels are usually complementary, with each channel telling a different part of the story. Curren…
Automatic Understanding of Image and Video Advertisements
Zaeem Hussain, Mingda Zhang, Xiaozhong Zhang +5
There is more to images than their objective physical content: for example, advertisements are created to persuade a viewer to take a certain action. We propose the novel problem o…