4 papers
Scal3R: Learning Efficient Multi-Relative Pose Query for Scalable Online 3D Reconstruction
Chin-Yang Lin, Yang-Che Sun, Cheng Sun +5
Online 3D reconstruction models perform poorly on long videos. This happens because regressing poses relative to a fixed first-frame anchor forces extrapolation far beyond the trai…
3AM: 3egment Anything with Geometric Consistency in Videos
Yang-Che Sun, Cheng Sun, Chin-Yang Lin +4
Video object segmentation methods like SAM2 achieve strong performance through memory-based architectures but struggle under large viewpoint changes due to reliance on appearance f…
FIPER: Factorized Features for Robust Image Super-Resolution and Compression
Yang-Che Sun, Cheng Yu Yeo, Ernie Chu +2
In this work, we propose using a unified representation, termed Factorized Features, for low-level vision tasks, where we test on Single Image Super-Resolution (SISR) and \textbf{I…
Task-Customized Mixture of Adapters for General Image Fusion
Pengfei Zhu, Yang Sun, Bing Cao +1
General image fusion aims at integrating important information from multi-source images. However, due to the significant cross-task gap, the respective fusion mechanism varies cons…