59 citations · 175 across the 31 of their papers we have counts for
18 papers · 1 filter
Dual Adaptive Representation Alignment for Cross-domain Few-shot Learning
Yifan Zhao, Tong Zhang, Jia Li +1
Few-shot learning aims to recognize novel queries with limited support samples by learning from base knowledge. Recent progress in this setting assumes that the base knowledge and…
ADAPT: Action-aware Driving Caption Transformer
Bu Jin, Xinyu Liu, Yupeng Zheng +6
End-to-end autonomous driving has great potential in the transportation industry. However, the lack of transparency and interpretability of the automatic decision-making process hi…
FAF: A novel multimodal emotion recognition approach integrating face, body and text
Zhongyu Fang, Aoyun He, Qihui Yu +4
Multimodal emotion analysis performed better in emotion recognition depending on more comprehensive emotional clues and multimodal emotion dataset. In this paper, we developed a la…
Involution: Inverting the Inherence of Convolution for Visual Recognition
Duo Li, Jie Hu, Changhu Wang +5
Convolution has been the core ingredient of modern neural networks, triggering the surge of deep learning in vision. In this work, we rethink the inherent principles of standard co…
Sparse Coding on Cascaded Residuals
Tong Zhang, Fatih Porikli
This paper seeks to combine dictionary learning and hierarchical image representation in a principled way. To make dictionary atoms capturing additional information from extended r…
Tencent ML-Images: A Large-Scale Multi-Label Image Database for Visual Representation Learning
Baoyuan Wu, Weidong Chen, Yanbo Fan +4
In existing visual representation learning tasks, deep convolutional neural networks (CNNs) are often trained on images annotated with single tags, such as ImageNet. However, a sin…