2 papers
cs.CV2026
SMFormer: Empowering Self-supervised Stereo Matching via Foundation Models and Data Augmentation
Yun Wang, Zhengjie Yang, Jiahao Zheng +3
Recent self-supervised stereo matching methods have made significant progress. They typically rely on the photometric consistency assumption, which presumes corresponding points ac…
cs.CV2025
Video-EM: Event-Centric Episodic Memory for Long-Form Video Understanding
Yun Wang, Long Zhang, Jingren Liu +9
Video Large Language Models (Video-LLMs) have shown strong video understanding, yet their application to long-form videos remains constrained by limited context windows. A common w…