2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CV2024★ 2 cited
Pretraining Vision-Language Model for Difference Visual Question Answering in Longitudinal Chest X-rays
Yeongjae Cho, Taehee Kim, Heejun Shin +2
Difference visual question answering (diff-VQA) is a challenging task that requires answering complex questions based on differences between a pair of images. This task is particul…
cs.CL2024
Generalizing Visual Question Answering from Synthetic to Human-Written Questions via a Chain of QA with a Large Language Model
Taehee Kim, Yeongjae Cho, Heejun Shin +2
Visual question answering (VQA) is a task where an image is given, and a series of questions are asked about the image. To build an efficient VQA algorithm, a large amount of QA da…