3 papers
physics.optics2026
Physically Grounded Monocular Depth via Nanophotonic Wavefront Encoding
Bingxuan Li, Jiahao Wu, Yuan Xu +6
Depth foundation models (DFMs) offer strong learned priors for 3D perception from single RGB images but lack physical depth cues, leading to ambiguities in metric scale. We introdu…
cs.CV2025
Image-GS: Content-Adaptive Image Representation via 2D Gaussians
Yunxiang Zhang, Bingxuan Li, Alexandr Kuznetsov +6
Neural image representations have emerged as a promising approach for encoding and rendering visual data. Combined with learning-based workflows, they demonstrate impressive trade-…
cs.CV2025
GazeFusion: Saliency-Guided Image Generation
Yunxiang Zhang, Nan Wu, Connor Z. Lin +2
Diffusion models offer unprecedented image generation power given just a text prompt. While emerging approaches for controlling diffusion models have enabled users to specify the d…