activity
20242026
collaborators

8 papers

cs.CV2026

Multi-Order Matching Network for Alignment-Free Depth Super-Resolution

Zhengxue Wang, Zhiqiang Yan, Yuan Wu +3

Recent guided depth super-resolution methods are premised on the assumption of strict spatial alignment between depth and RGB, achieving high-quality depth reconstruction. However,…

cs.CV2025

SpatioTemporal Difference Network for Video Depth Super-Resolution

Zhengxue Wang, Yuan Wu, Xiang Li +2

Depth super-resolution has achieved impressive performance, and the incorporation of multi-frame information further enhances reconstruction quality. Nevertheless, statistical anal…

cs.CV2025

See through the Dark: Learning Illumination-affined Representations for Nighttime Occupancy Prediction

Yuan Wu, Zhiqiang Yan, Yigong Zhang +2

Occupancy prediction aims to estimate the 3D spatial distribution of occupied regions along with their corresponding semantic labels. Existing vision-based methods perform well on…

cs.DL2025

A Literature Review of Literature Reviews in Pattern Analysis and Machine Intelligence

Penghai Zhao, Xin Zhang, Jiayue Cao +3

The rapid growth of research in Pattern Analysis and Machine Intelligence (PAMI) has rendered literature reviews essential for consolidating and interpreting knowledge across its m…

cs.CV2025

Deep Height Decoupling for Precise Vision-based 3D Occupancy Prediction

Yuan Wu, Zhiqiang Yan, Zhengxue Wang +3

The task of vision-based 3D occupancy prediction aims to reconstruct 3D geometry and estimate its semantic classes from 2D color images, where the 2D-to-3D view transformation is a…

cs.CV2025

OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation

Kepan Nan, Rui Xie, Penghao Zhou +6

Text-to-video (T2V) generation has recently garnered significant attention thanks to the large multi-modality model Sora. However, T2V generation still faces two important challeng…