collaborators

8 papers

cs.RO2026

Global-Local Attention Decomposition for Terrain Encoding in Humanoid Perceptive Locomotion

Shengcheng Fu, Yang Zhang, Zhanxiang Cao +4

Although reinforcement learning has significantly advanced humanoid locomotion, perceptive policies still struggle on sparse-foothold terrain and constrained environments. Success…

cs.CV2026

αDepth: Learning Single-Pass Soft Boundary Decomposition for Stereo Conversion

Xiang Zhang, Yang Zhang, Lukas Mehl +3

Accurately modeling soft boundaries, e.g., hair and defocus blur, is a fundamental challenge in stereo conversion due to the ambiguous blending of foreground and background. Existi…

cs.CV2026

UniFixer: A Universal Reference-Guided Fixer for Diffusion-Based View Synthesis

Sihan Chen, Xiang Zhang, Yang Zhang +2

With the recent surge of generative models, diffusion-based approaches have become mainstream for view synthesis tasks, either in an explicit depth-warp-inpaint or in an implicit e…

cs.CV2026

RenderFlow: Single-Step Neural Rendering via Flow Matching

Shenghao Zhang, Runtao Liu, Christopher Schroers +1

Conventional physically based rendering (PBR) pipelines generate photorealistic images through computationally intensive light transport simulations. Although recent deep learning…

eess.IV2026

Region-Adaptive Generative Compression with Spatially Varying Diffusion Models

Lucas Relic, Roberto Azevedo, Yang Zhang +3

Generative image codecs aim to optimize perceptual quality, producing realistic and detailed reconstructions. However, they often overlook a key property of human vision: our tende…

cs.CV2026

Guardians of the Hair: Rescuing Soft Boundaries in Depth, Stereo, and Novel Views

Xiang Zhang, Yang Zhang, Lukas Mehl +2

Soft boundaries, like thin hairs, are commonly observed in natural and computer-generated imagery, but they remain challenging for 3D vision due to the ambiguous mixing of foregrou…