6 citations · 12 across the 4 of their papers we have counts for
1 paper · 1 filter
Soyeon Caren Han, Siqu Long, Siwen Luo +2
Text-to-image multimodal tasks, generating/retrieving an image from a given text description, are extremely challenging tasks since raw text descriptions cover quite limited inform…