6 citations · 7 across the 2 of their papers we have counts for
2 papers
eess.IV2024★ 6 cited
Window-based Channel Attention for Wavelet-enhanced Learned Image Compression
Heng Xu, Bowen Hai, Yushun Tang +1
Learned Image Compression (LIC) models have achieved superior rate-distortion performance than traditional codecs. Existing LIC models use CNN, Transformer, or Mixed CNN-Transforme…
cs.CV2023★ 1 cited
Learning to Adapt CLIP for Few-Shot Monocular Depth Estimation
Xueting Hu, Ce Zhang, Yi Zhang +3
Pre-trained Vision-Language Models (VLMs), such as CLIP, have shown enhanced performance across a range of tasks that involve the integration of visual and linguistic modalities. W…