1 citations · 2 across the 4 of their papers we have counts for
4 papers
Attention-guided Multi-step Fusion: A Hierarchical Fusion Network for Multimodal Recommendation
Yan Zhou, Jie Guo, Hao Sun +2
The main idea of multimodal recommendation is the rational utilization of the item's multimodal information to improve the recommendation performance. Previous works directly integ…
MVP-SEG: Multi-View Prompt Learning for Open-Vocabulary Semantic Segmentation
Jie Guo, Qimeng Wang, Yan Gao +4
CLIP (Contrastive Language-Image Pretraining) is well-developed for open-vocabulary zero-shot image-level recognition, while its applications in pixel-level tasks are less investig…
Local-to-Global Panorama Inpainting for Locale-Aware Indoor Lighting Prediction
Jiayang Bai, Zhen He, Shan Yang +4
Predicting panoramic indoor lighting from a single perspective image is a fundamental but highly ill-posed problem in computer vision and graphics. To achieve locale-aware and robu…
Self-NeRF: A Self-Training Pipeline for Few-Shot Neural Radiance Fields
Jiayang Bai, Letian Huang, Wen Gong +2
Recently, Neural Radiance Fields (NeRF) have emerged as a potent method for synthesizing novel views from a dense set of images. Despite its impressive performance, NeRF is plagued…