3 papers
cs.IR2024
Cross-Modal Retrieval: A Systematic Review of Methods and Future Directions
Tianshi Wang, Fengling Li, Lei Zhu +3
With the exponential surge in diverse multi-modal data, traditional uni-modal retrieval methods struggle to meet the needs of users seeking access to data across various modalities…
cs.CV2024
Caterpillar: A Pure-MLP Architecture with Shifted-Pillars-Concatenation
Jin Sun, Xiaoshuang Shi, Zhiyuan Wang +3
Modeling in Computer Vision has evolved to MLPs. Vision MLPs naturally lack local modeling capability, to which the simplest treatment is combined with convolutional layers. Convol…
cs.CV2024
JoReS-Diff: Joint Retinex and Semantic Priors in Diffusion Model for Low-light Image Enhancement
Yuhui Wu, Guoqing Wang, Zhiwen Wang +5
Low-light image enhancement (LLIE) has achieved promising performance by employing conditional diffusion models. Despite the success of some conditional methods, previous methods m…