6 citations · 16 across the 4 of their papers we have counts for
5 papers
An Improved Attention for Visual Question Answering
Tanzila Rahman, Shih-Han Chou, Leonid Sigal +1
We consider the problem of Visual Question Answering (VQA). Given an image and a free-form, open-ended, question, expressed in natural language, the goal of VQA system is to provid…
Visual Question Answering on 360° Images
Shih-Han Chou, Wei-Lun Chao, Wei-Sheng Lai +2
In this work, we introduce VQA 360, a novel task of visual question answering on 360 images. Unlike a normal field-of-view image, a 360 image captures the entire visual content aro…
360-Indoor: Towards Learning Real-World Objects in 360° Indoor Equirectangular Images
Shih-Han Chou, Cheng Sun, Wen-Yen Chang +3
While there are several widely used object detection datasets, current computer vision algorithms are still limited in conventional images. Such images narrow our vision in a restr…
Self-view Grounding Given a Narrated 360° Video
Shih-Han Chou, Yi-Chun Chen, Kuo-Hao Zeng +3
Narrated 360° videos are typically provided in many touring scenarios to mimic real-world experience. However, previous work has shown that smart assistance (i.e., providing visual…
Agent-Centric Risk Assessment: Accident Anticipation and Risky Region Localization
Kuo-Hao Zeng, Shih-Han Chou, Fu-Hsiang Chan +2
For survival, a living agent must have the ability to assess risk (1) by temporally anticipating accidents before they occur, and (2) by spatially localizing risky regions in the e…