2 papers
cs.CV2025
CLIPin: A Non-contrastive Plug-in to CLIP for Multimodal Semantic Alignment
Shengzhu Yang, Jiawei Du, Shuai Lu +3
Large-scale natural image-text datasets, especially those automatically collected from the web, often suffer from loose semantic alignment due to weak supervision, while medical da…
cs.CV2024
PVP: Polar Representation Boost for 3D Semantic Occupancy Prediction
Yujing Xue, Jiaxiang Liu, Jiawei Du +1
Recently, polar coordinate-based representations have shown promise for 3D perceptual tasks. Compared to Cartesian methods, polar grids provide a viable alternative, offering bette…