3 papers
cs.CL2024
Modularized Networks for Few-shot Hateful Meme Detection
Rui Cao, Roy Ka-Wei Lee, Jing Jiang
In this paper, we address the challenge of detecting hateful memes in the low-resource setting where only a few labeled examples are available. Our approach leverages the compositi…
cs.CL2024
Knowledge Generation for Zero-shot Knowledge-based VQA
Rui Cao, Jing Jiang
Previous solutions to knowledge-based visual question answering~(K-VQA) retrieve knowledge from external knowledge bases and use supervised learning to train the K-VQA model. Recen…
cs.CV2024
Modularized Zero-shot VQA with Pre-trained Models
Rui Cao, Jing Jiang
Large-scale pre-trained models (PTMs) show great zero-shot capabilities. In this paper, we study how to leverage them for zero-shot visual question answering (VQA). Our approach is…