1 citations · 1 across the 1 of their papers we have counts for
Showing cs.CVShow all
2 papers · 1 filter
cs.CV2023
UIT-Saviors at MEDVQA-GI 2023: Improving Multimodal Learning with Image Enhancement for Gastrointestinal Visual Question Answering
Triet M. Thai, Anh T. Vo, Hao K. Tieu +2
In recent years, artificial intelligence has played an important role in medicine and disease diagnosis, with many applications to be mentioned, one of which is Medical Visual Ques…
cs.CV2023
Integrating Image Features with Convolutional Sequence-to-sequence Network for Multilingual Visual Question Answering
Triet Minh Thai, Son T. Luu
Visual Question Answering (VQA) is a task that requires computers to give correct answers for the input questions based on the images. This task can be solved by humans with ease b…