1 paper
Harish Haresamudram, Apoorva Beedu, Mashfiqui Rabbi +3
Cross-modal contrastive pre-training between natural language and other modalities, e.g., vision and audio, has demonstrated astonishing performance and effectiveness across a dive…