8 papers
Global-Local Attention Decomposition for Terrain Encoding in Humanoid Perceptive Locomotion
Shengcheng Fu, Yang Zhang, Zhanxiang Cao +4
Although reinforcement learning has significantly advanced humanoid locomotion, perceptive policies still struggle on sparse-foothold terrain and constrained environments. Success…
αDepth: Learning Single-Pass Soft Boundary Decomposition for Stereo Conversion
Xiang Zhang, Yang Zhang, Lukas Mehl +3
Accurately modeling soft boundaries, e.g., hair and defocus blur, is a fundamental challenge in stereo conversion due to the ambiguous blending of foreground and background. Existi…
UniFixer: A Universal Reference-Guided Fixer for Diffusion-Based View Synthesis
Sihan Chen, Xiang Zhang, Yang Zhang +2
With the recent surge of generative models, diffusion-based approaches have become mainstream for view synthesis tasks, either in an explicit depth-warp-inpaint or in an implicit e…
RenderFlow: Single-Step Neural Rendering via Flow Matching
Shenghao Zhang, Runtao Liu, Christopher Schroers +1
Conventional physically based rendering (PBR) pipelines generate photorealistic images through computationally intensive light transport simulations. Although recent deep learning…
Region-Adaptive Generative Compression with Spatially Varying Diffusion Models
Lucas Relic, Roberto Azevedo, Yang Zhang +3
Generative image codecs aim to optimize perceptual quality, producing realistic and detailed reconstructions. However, they often overlook a key property of human vision: our tende…
Guardians of the Hair: Rescuing Soft Boundaries in Depth, Stereo, and Novel Views
Xiang Zhang, Yang Zhang, Lukas Mehl +2
Soft boundaries, like thin hairs, are commonly observed in natural and computer-generated imagery, but they remain challenging for 3D vision due to the ambiguous mixing of foregrou…