5 papers
Hyperbolic Hierarchical Alignment Reasoning Network for Text-3D Retrieval
Wenrui Li, Yidan Lu, Yeyu Chai +3
With the daily influx of 3D data on the internet, text-3D retrieval has gained increasing attention. However, current methods face two major challenges: Hierarchy Representation Co…
Language-Guided Graph Representation Learning for Video Summarization
Wenrui Li, Wei Han, Hengyu Man +3
With the rapid growth of video content on social media, video summarization has become a crucial task in multimedia processing. However, existing methods face challenges in capturi…
MRT: Learning Compact Representations with Mixed RWKV-Transformer for Extreme Image Compression
Han Liu, Hengyu Man, Xingtao Wang +2
Recent advances in extreme image compression have revealed that mapping pixel data into highly compact latent representations can significantly improve coding efficiency. However,…
T-GVC: Trajectory-Guided Generative Video Coding at Ultra-Low Bitrates
Zhitao Wang, Hengyu Man, Wenrui Li +3
Recent advances in video generation techniques have given rise to an emerging paradigm of generative video coding for Ultra-Low Bitrate (ULB) scenarios by leveraging powerful gener…
Hyperbolic-constraint Point Cloud Reconstruction from Single RGB-D Images
Wenrui Li, Zhe Yang, Wei Han +3
Reconstructing desired objects and scenes has long been a primary goal in 3D computer vision. Single-view point cloud reconstruction has become a popular technique due to its low c…