activity
20162024
most citedInjecting Semantic Concepts into End-to-End Image Captioning

8 citations · 14 across the 12 of their papers we have counts for

collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV20218 cited

Injecting Semantic Concepts into End-to-End Image Captioning

Zhiyuan Fang, Jianfeng Wang, Xiaowei Hu +5

Tremendous progress has been made in recent years in developing better image captioning models, yet most of them rely on a separate object detector to extract regional features. Re…

cs.CV2016

Fast Task-Specific Target Detection via Graph Based Constraints Representation and Checking

Went Luan, Yezhou Yang, Cornelia Fermuller +1

In this work, we present a fast target detection framework for real-world robotics applications. Considering that an intelligent agent attends to a task-specific object target duri…

cs.CV20163 cited

Answering Image Riddles using Vision and Reasoning through Probabilistic Soft Logic

Somak Aditya, Yezhou Yang, Chitta Baral +1

In this work, we explore a genre of puzzles ("image riddles") which involves a set of images and a question. Answering these puzzles require both capabilities involving visual dete…

cs.CV2016

Reliable Attribute-Based Object Recognition Using High Predictive Value Classifiers

Wentao Luan, Yezhou Yang, Cornelia Fermuller +1

We consider the problem of object recognition in 3D using an ensemble of attribute-based classifiers. We propose two new concepts to improve classification in practical situations,…

cs.CV20162 cited

Prediction of Manipulation Actions

Cornelia Fermüller, Fang Wang, Yezhou Yang +4

Looking at a person's hands one often can tell what the person is going to do next, how his/her hands are moving and where they will be, because an actor's intentions shape his/her…