48 citations · 48 across the 4 of their papers we have counts for
4 papers
CLIP2GAN: Towards Bridging Text with the Latent Space of GANs
Yixuan Wang, Wengang Zhou, Jianmin Bao +3
In this work, we are dedicated to text-guided image generation and propose a novel framework, i.e., CLIP2GAN, by leveraging CLIP model and StyleGAN. The key idea of our CLIP2GAN is…
Music Similarity Calculation of Individual Instrumental Sounds Using Metric Learning
Yuka Hashizume, Li Li, Tomoki Toda
The criteria for measuring music similarity are important for developing a flexible music recommendation system. Some data-driven methods have been proposed to calculate music simi…
DeepMLE: A Robust Deep Maximum Likelihood Estimator for Two-view Structure from Motion
Yuxi Xiao, Li Li, Xiaodi Li +1
Two-view structure from motion (SfM) is the cornerstone of 3D reconstruction and visual SLAM (vSLAM). Many existing end-to-end learning-based methods usually formulate it as a brut…
Attribute Artifacts Removal for Geometry-based Point Cloud Compression
Xihua Sheng, Li Li, Dong Liu +1
Geometry-based point cloud compression (G-PCC) can achieve remarkable compression efficiency for point clouds. However, it still leads to serious attribute compression artifacts, e…