13 citations · 26 across the 3 of their papers we have counts for
1 paper · 1 filter
Mingchen Zhuge, Dehong Gao, Deng-Ping Fan +5
We present a new vision-language (VL) pre-training model dubbed Kaleido-BERT, which introduces a novel kaleido strategy for fashion cross-modality representations from transformers…