3 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.CV2026
Text-Guided Visual Dependency Graph Learning with Cross-Modal Attention Priors
Fei Wang, Yutong Zhang, Yang Ye +3
Estimating interpretable conditional-dependence structures from multimodal visual-linguistic features remains largely unexplored. We propose CM-GLasso (Cross-Modal Graphical Lasso)…
cs.CV2026
Multimodal Structure Learning: Disentangling Shared and Specific Topology via Cross-Modal Graphical Lasso
Fei Wang, Yutong Zhang, Xiong Wang
Learning interpretable multimodal representations inherently relies on uncovering the conditional dependencies between heterogeneous features. However, sparse graph estimation tech…
cs.CV2025★ 3 cited
SMART: Semantic Matching Contrastive Learning for Partially View-Aligned Clustering
Liang Peng, Yixuan Ye, Cheng Liu +5
Multi-view clustering has been empirically shown to improve learning performance by leveraging the inherent complementary information across multiple views of data. However, in rea…