24 citations · 46 across the 5 of their papers we have counts for
1 paper · 1 filter
Elyas Meguellati, Nardiena Pratama, Shazia Sadiq +1
High-quality textual training data is essential for the success of multimodal data processing tasks, yet outputs from image captioning models like BLIP and GIT often contain errors…