3 papers
cs.LG2025
Learning from Ambiguous Data with Hard Labels
Zeke Xie, Zheng He, Nan Lu +5
Real-world data often contains intrinsic ambiguity that the common single-hard-label annotation paradigm ignores. Standard training using ambiguous data with these hard labels may…
cs.CV2024
VIP: Versatile Image Outpainting Empowered by Multimodal Large Language Model
Jinze Yang, Haoran Wang, Zining Zhu +3
In this paper, we focus on resolving the problem of image outpainting, which aims to extrapolate the surrounding parts given the center contents of an image. Although recent works…
cs.CV2024
SGD: Street View Synthesis with Gaussian Splatting and Diffusion Prior
Zhongrui Yu, Haoran Wang, Jinze Yang +6
Novel View Synthesis (NVS) for street scenes play a critical role in the autonomous driving simulation. The current mainstream technique to achieve it is neural rendering, such as…