Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Understanding Multi-Agent Reasoning with Large Language Models for Cartoon VQA
Tong Wu, Thanet Markchom
Visual Question Answering (VQA) for stylised cartoon imagery presents challenges, such as interpreting exaggerated visual abstraction and narrative-driven context, which are not ad…
cs.CV2024
CRRG-CLIP: Automatic Generation of Chest Radiology Reports and Classification of Chest Radiographs
Jianfei Xu, Thanet Markchom, Huizhi Liang
The complexity of stacked imaging and the massive number of radiographs make writing radiology reports complex and inefficient. Even highly experienced radiologists struggle to mai…