8 citations · 13 across the 6 of their papers we have counts for
6 papers
CAROM Air -- Vehicle Localization and Traffic Scene Reconstruction from Aerial Videos
Duo Lu, Eric Eaton, Matt Weg +5
Road traffic scene reconstruction from videos has been desirable by road safety regulators, city planners, researchers, and autonomous driving technology developers. However, it is…
Injecting Semantic Concepts into End-to-End Image Captioning
Zhiyuan Fang, Jianfeng Wang, Xiaowei Hu +5
Tremendous progress has been made in recent years in developing better image captioning models, yet most of them rely on a separate object detector to extract regional features. Re…
Fast Task-Specific Target Detection via Graph Based Constraints Representation and Checking
Went Luan, Yezhou Yang, Cornelia Fermuller +1
In this work, we present a fast target detection framework for real-world robotics applications. Considering that an intelligent agent attends to a task-specific object target duri…
Answering Image Riddles using Vision and Reasoning through Probabilistic Soft Logic
Somak Aditya, Yezhou Yang, Chitta Baral +1
In this work, we explore a genre of puzzles ("image riddles") which involves a set of images and a question. Answering these puzzles require both capabilities involving visual dete…
Reliable Attribute-Based Object Recognition Using High Predictive Value Classifiers
Wentao Luan, Yezhou Yang, Cornelia Fermuller +1
We consider the problem of object recognition in 3D using an ensemble of attribute-based classifiers. We propose two new concepts to improve classification in practical situations,…
Prediction of Manipulation Actions
Cornelia Fermüller, Fang Wang, Yezhou Yang +4
Looking at a person's hands one often can tell what the person is going to do next, how his/her hands are moving and where they will be, because an actor's intentions shape his/her…