6 citations · 9 across the 29 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2021
Collaborative Learning to Generate Audio-Video Jointly
Vinod K Kurmi, Vipul Bajaj, Badri N Patro +3
There have been a number of techniques that have demonstrated the generation of multimedia data for one modality at a time using GANs, such as the ability to generate images, video…
cs.CV2021
Select, Substitute, Search: A New Benchmark for Knowledge-Augmented Visual Question Answering
Aman Jain, Mayank Kothyari, Vishwajeet Kumar +3
Multimodal IR, spanning text corpus, knowledge graph and images, called outside knowledge visual question answering (OKVQA), is of much recent interest. However, the popular data s…