3 papers
cs.CV2026
MambaPanoptic: A Vision Mamba-based Structured State Space Framework for Panoptic Segmentation
Qing Cheng, Damiano Bertolini, Wei Zhang +3
Panoptic segmentation requires the simultaneous recognition of countable thing instances and amorphous stuff regions, placing joint demands on long-range context modelling, multi-s…
cs.CV2025
When and Where do Events Switch in Multi-Event Video Generation?
Ruotong Liao, Guowen Huang, Qing Cheng +3
Text-to-video (T2V) generation has surged in response to challenging questions, especially when a long video must depict multiple sequential events with temporal coherence and cont…
cs.CV2025
PRISM: Probabilistic Representation for Integrated Shape Modeling and Generation
Lei Cheng, Mahdi Saleh, Qing Cheng +4
Despite the advancements in 3D full-shape generation, accurately modeling complex geometries and semantics of shape parts remains a significant challenge, particularly for shapes w…