Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Unleashing the Multi-View Fusion Potential: Noise Correction in VLM for Open-Vocabulary 3D Scene Understanding
Xingyilang Yin, Jiale Wang, Xi Yang +3
Recent open-vocabulary 3D scene understanding approaches mainly focus on training 3D networks through contrastive learning with point-text pairs or by distilling 2D features into 3…
cs.CV2025
Pro2SAM: Mask Prompt to SAM with Grid Points for Weakly Supervised Object Localization
Xi Yang, Songsong Duan, Nannan Wang +1
Weakly Supervised Object Localization (WSOL), which aims to localize objects by only using image-level labels, has attracted much attention because of its low annotation cost in re…
cs.CV2025
CSHNet: A Novel Information Asymmetric Image Translation Method
Xi Yang, Haoyuan Shi, Zihan Wang +2
Despite advancements in cross-domain image translation, challenges persist in asymmetric tasks such as SAR-to-Optical and Sketch-to-Instance conversions, which involve transforming…