3 papers
cs.LG2026
Learning Multimodal Energy-Based Model with Multimodal Variational Auto-Encoder via MCMC Revision
Jiali Cui, Zhiqiang Lao, Heather Yu
Energy-based models (EBMs) are a flexible class of deep generative models and are well-suited to capture complex dependencies in multimodal data. However, learning multimodal EBM b…
cs.GR2025
Inverse Rendering for High-Genus Surface Meshes from Multi-View Images
Xiang Gao, Xinmu Wang, Xiaolong Wu +8
We present a topology-informed inverse rendering approach for reconstructing high-genus surface meshes from multi-view images. Compared to 3D representations like voxels and point…
cs.CV2025
Scene Perceived Image Perceptual Score (SPIPS): combining global and local perception for image quality assessment
Zhiqiang Lao, Heather Yu
The rapid advancement of artificial intelligence and widespread use of smartphones have resulted in an exponential growth of image data, both real (camera-captured) and virtual (AI…