Neural Inverse Rendering for General Reflectance Photometric Stereo
arXiv:1802.10328
Abstract
We present a novel convolutional neural network architecture for photometric stereo (Woodham, 1980), a problem of recovering 3D object surface normals from multiple images observed under varying illuminations. Despite its long history in computer vision, the problem still shows fundamental challenges for surfaces with unknown general reflectance properties (BRDFs). Leveraging deep neural networks to learn complicated reflectance models is promising, but studies in this direction are very limited due to difficulties in acquiring accurate ground truth for training and also in designing networks invariant to permutation of input images. In order to address these challenges, we propose a physics based unsupervised learning framework where surface normals and BRDFs are predicted by the network and fed into the rendering equation to synthesize observed images. The network weights are optimized during testing by minimizing reconstruction loss between observed and synthesized images. Thus, our learning process does not require ground truth normals or even pre-training on external images. Our method is shown to achieve the state-of-the-art performance on a challenging real-world scene benchmark.
To appear in International Conference on Machine Learning 2018 (ICML 2018). 10 pages + 20 pages (appendices)
Cited by in corpus (15)
- RenderNet: A deep convolutional network for differentiable rendering from 3D shapes
- DeProCams: Simultaneous Relighting, Compensation and Shape Reconstruction for Projector-Camera Systems
- Learning Inter- and Intraframe Representations for Non-Lambertian Photometric Stereo
- A CNN Based Approach for the Point-Light Photometric Stereo Problem
- Single Day Outdoor Photometric Stereo
- SPLINE-Net: Sparse Photometric Stereo through Lighting Interpolation and Normal Estimation Networks
- Self-calibrating Deep Photometric Stereo Networks
- Deep Shape from Polarization
- MS-PS: A Multi-Scale Network for Photometric Stereo With a New Comprehensive Training Dataset
- Deep Photometric Stereo for Non-Lambertian Surfaces
- RMAFF-PSN: A Residual Multi-Scale Attention Feature Fusion Photometric Stereo Network
- Lightweight Photometric Stereo for Facial Details Recovery
- Image Gradient-Aided Photometric Stereo Network
- IntrinSeqNet: Learning to Estimate the Reflectance from Varying Illumination
- A Dark Flash Normal Camera