84 citations · 189 across the 13 of their papers we have counts for
13 papers · 1 filter
CMA-CLIP: Cross-Modality Attention CLIP for Image-Text Classification
Huidong Liu, Shaoyuan Xu, Jinmiao Fu +5
Modern Web systems such as social media and e-commerce contain rich contents expressed in images and text. Leveraging information from multi-modalities can improve the performance…
TransZero: Attribute-guided Transformer for Zero-Shot Learning
Shiming Chen, Ziming Hong, Yang Liu +6
Zero-shot learning (ZSL) aims to recognize novel classes by transferring semantic knowledge from seen classes to unseen ones. Semantic knowledge is learned from attribute descripti…
Cross Modal Retrieval with Querybank Normalisation
Simion-Vlad Bogolin, Ioana Croitoru, Hailin Jin +2
Profiting from large-scale training datasets, advances in neural architecture design and efficient inference, joint embeddings have become the dominant approach for tackling cross-…
TCGL: Temporal Contrastive Graph for Self-supervised Video Representation Learning
Yang Liu, Keze Wang, Lingbo Liu +2
Video self-supervised learning is a challenging task, which requires significant expressive power from the model to leverage rich spatial-temporal knowledge and generate effective…
Causal Homotopy
Sridhar Mahadevan
We characterize homotopical equivalences between causal DAG models, exploiting the close connections between partially ordered set representations of DAGs (posets) and finite Alexa…
Deconvolutional Networks on Graph Data
Jia Li, Jiajin Li, Yang Liu +3
In this paper, we consider an inverse problem in graph learning domain -- ``given the graph representations smoothed by Graph Convolutional Network (GCN), how can we reconstruct th…