7 citations · 13 across the 7 of their papers we have counts for
9 papers · 1 filter
The Sound of Bounding-Boxes
Takashi Oya, Shohei Iwase, Shigeo Morishima
In the task of audio-visual sound source separation, which leverages visual information for sound source separation, identifying objects in an image is a crucial step prior to sepa…
360 Depth Estimation in the Wild -- The Depth360 Dataset and the SegFuse Network
Qi Feng, Hubert P. H. Shum, Shigeo Morishima
Single-view depth estimation from omnidirectional images has gained popularity with its wide range of applications such as autonomous driving and scene reconstruction. Although dat…
Do We Need Sound for Sound Source Localization?
Takashi Oya, Shohei Iwase, Ryota Natsume +3
During the performance of sound source localization which uses both visual and aural information, it presently remains unclear how much either image or sound modalities contribute…
MirrorNet: A Deep Bayesian Approach to Reflective 2D Pose Estimation from Human Images
Takayuki Nakatsuka, Kazuyoshi Yoshii, Yuki Koyama +3
This paper proposes a statistical approach to 2D pose estimation from human images. The main problems with the standard supervised approach, which is based on a deep recognition (i…
What Do Adversarially Robust Models Look At?
Takahiro Itazuri, Yoshihiro Fukuhara, Hirokatsu Kataoka +1
In this paper, we address the open question: "What do adversarially robust models look at?" Recently, it has been reported in many works that there exists the trade-off between sta…
PIFu: Pixel-Aligned Implicit Function for High-Resolution Clothed Human Digitization
Shunsuke Saito, Zeng Huang, Ryota Natsume +3
We introduce Pixel-aligned Implicit Function (PIFu), a highly effective implicit representation that locally aligns pixels of 2D images with the global context of their correspondi…