3 papers
cs.CV2026
A General-Purpose VLM Can Teach an Astronomy Foundation Model to Better Recognize Galaxy Morphology
Dichang Zhang, Jiaqi Deng, Yixuan Shao +7
Existing astronomy foundation models provide strong galaxy representations, but adapting them to new survey conditions and survey-specific morphology recognition tasks still requir…
cs.LG2026
Learning Multimodal Energy-Based Model with Multimodal Variational Auto-Encoder via MCMC Revision
Jiali Cui, Zhiqiang Lao, Heather Yu
Energy-based models (EBMs) are a flexible class of deep generative models and are well-suited to capture complex dependencies in multimodal data. However, learning multimodal EBM b…
cs.LG2025
ShaLa: Multimodal Shared Latent Space Modelling
Jiali Cui, Yan-Ying Chen, Yanxia Zhang +1
This paper presents a novel generative framework for learning shared latent representations across multimodal data. Many advanced multimodal methods focus on capturing all combinat…