3 papers
cs.CV2026
OrbitNVS: Harnessing Video Diffusion Priors for Novel View Synthesis
Jinglin Liang, Zijian Zhou, Rui Huang +2
Novel View Synthesis (NVS) aims to generate unseen views of a 3D object given a limited number of known views. Existing methods often struggle to synthesize plausible views for uno…
cs.CV2025
MPDrive: Improving Spatial Understanding with Marker-Based Prompt Learning for Autonomous Driving
Zhiyuan Zhang, Xiaofan Li, Zhihao Xu +4
Autonomous driving visual question answering (AD-VQA) aims to answer questions related to perception, prediction, and planning based on given driving scene images, heavily relying…
cs.CV2024
SEG-SAM: Semantic-Guided SAM for Unified Medical Image Segmentation
Shuangping Huang, Hao Liang, Qingfeng Wang +3
Recently, developing unified medical image segmentation models gains increasing attention, especially with the advent of the Segment Anything Model (SAM). SAM has shown promising b…