2 papers
cs.AI2026
IMUG-Bench: Benchmarking Unified Multimodal Models on Interleaved Understanding and Generation
Lingyi Meng, Zecong Tang, Haoran Li +12
In recent years, unified multimodal models (UMMs) have emerged to support both understanding and generation within a single framework. Mastering dynamic, multi-turn interleaved ima…
cs.CV2025
PROMISE: Prompt-Attentive Hierarchical Contrastive Learning for Robust Cross-Modal Representation with Missing Modalities
Jiajun Chen, Sai Cheng, Yutao Yuan +4
Multimodal models integrating natural language and visual information have substantially improved generalization of representation models. However, their effectiveness significantl…