Showing 2026Show all
2 papers · 1 filter
physics.optics2026
Physically Grounded Monocular Depth via Nanophotonic Wavefront Encoding
Bingxuan Li, Jiahao Wu, Yuan Xu +6
Depth foundation models (DFMs) offer strong learned priors for 3D perception from single RGB images but lack physical depth cues, leading to ambiguities in metric scale. We introdu…
cs.CV2026
Cost-Aware Routing for Efficient Text-To-Image Generation
Qinchan Li, Kenneth Chen, Changyue Su +3
Diffusion models are well known for their ability to generate a high-fidelity image for an input prompt through an iterative denoising process. Unfortunately, the high fidelity als…