Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Hub-Spectral Activation of Latent Multimodal Knowledge
Ying Guo, Haidong Chen, Linrui Xu +7
Multimodal representation learning seeks shared representations for cross-modal retrieval and knowledge transfer. Hub-based binding reduces pairwise supervision costs, but separate…
cs.CV2025
UniFace: A Unified Fine-grained Face Understanding and Generation Model
Junzhe Li, Sifan Zhou, Liya Guo +9
Unified multimodal models (UMMs) have emerged as a powerful paradigm in fundamental cross-modality research, demonstrating significant potential in both image understanding and gen…