5 papers
VSRD++: Autolabeling for 3D Object Detection via Instance-Aware Volumetric Silhouette Rendering
Zihua Liu, Hiroki Sakuma, Masatoshi Okutomi
Monocular 3D object detection is a fundamental yet challenging task in 3D scene understanding. Existing approaches heavily depend on supervised learning with extensive 3D annotatio…
DMS:Diffusion-Based Multi-Baseline Stereo Generation for Improving Self-Supervised Depth Estimation
Zihua Liu, Yizhou Li, Songyan Zhang +1
While supervised stereo matching and monocular depth estimation have advanced significantly with learning-based algorithms, self-supervised methods using stereo images as supervisi…
Segmentation-Guided Neural Radiance Fields for Novel Street View Synthesis
Yizhou Li, Yusuke Monno, Masatoshi Okutomi +3
Recent advances in Neural Radiance Fields (NeRF) have shown great potential in 3D reconstruction and novel view synthesis, particularly for indoor and small-scale scenes. However,…
TDM: Temporally-Consistent Diffusion Model for All-in-One Real-World Video Restoration
Yizhou Li, Zihua Liu, Yusuke Monno +1
In this paper, we propose the first diffusion-based all-in-one video restoration method that utilizes the power of a pre-trained Stable Diffusion and a fine-tuned ControlNet. Our m…
Disparity Estimation Using a Quad-Pixel Sensor
Zhuofeng Wu, Doehyung Lee, Zihua Liu +3
A quad-pixel (QP) sensor is increasingly integrated into commercial mobile cameras. The QP sensor has a unit of 22 four photodiodes under a single microlens, generating mul…