1 paper
Ethan Shen, Scotty Singh, Bhavesh Kumar
Multi-modal tasks involving vision and language in deep learning continue to rise in popularity and are leading to the development of newer models that can generalize beyond the ex…