Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
COMET: Contrastive Motion-Enhanced Temporal Reasoning for Video Multimodal Large Language Models
Chenghua Zhu, Zhaolu Kang, Qifan Shi +8
Video multimodal large language models have advanced significantly, yet fine-grained motion-temporal understanding remains fragile. The core bottleneck is not only sparse frame sam…
cs.CV2026
Zero-Forgetting CISS via Dual-Phase Cognitive Cascades
Yuquan Lu, Yifu Guo, Zishan Xu +6
Continual semantic segmentation (CSS) is a cornerstone task in computer vision that enables a large number of downstream applications, but faces the catastrophic forgetting challen…