2 papers
cs.CL2026
ManGo: Manga Active Narrative Grounding Optimization
Hao Qiu, Junyan Wang, Zheyuan Liu +4
Manga visual question answering requires models to answer questions over panel-based visual narratives, where relevant evidence is distributed across ordered panels, embedded text,…
cs.CV2026
Straight-Path Flow Matching for Incomplete Multi-View Clustering
Yiteng Yuan, Junyan Wang, Zheyuan Liu +4
Incomplete Multi-View Clustering addresses the problem of clustering multi-modal data when certain views are missing. Recent end-to-end generative approaches leverage diffusion mod…