3 papers
cs.CV2025
OpenView: Empowering MLLMs with Out-of-view VQA
Qixiang Chen, Cheng Zhang, Chi-Wing Fu +2
Recent multimodal large language models (MLLMs) show great potential in natural image understanding. Yet, they perform well, mainly on reasoning in-view contents within the image f…
cs.CV2025
ASAP-Textured Gaussians: Enhancing Textured Gaussians with Adaptive Sampling and Anisotropic Parameterization
Meng Wei, Cheng Zhang, Jianmin Zheng +2
Recent advances have equipped 3D Gaussian Splatting with texture parameterizations to capture spatially varying attributes, improving the performance of both appearance modeling an…
cs.CV2025
PanFlow: Decoupled Motion Control for Panoramic Video Generation
Cheng Zhang, Hanwen Liang, Donny Y. Chen +4
Panoramic video generation has attracted growing attention due to its applications in virtual reality and immersive media. However, existing methods lack explicit motion control an…