3 citations · 4 across the 3 of their papers we have counts for
3 papers
cs.CV2024★ 1 cited
TRRG: Towards Truthful Radiology Report Generation With Cross-modal Disease Clue Enhanced Large Language Model
Yuhao Wang, Chao Hao, Yawen Cui +4
The vision-language modeling capability of multi-modal large language models has attracted wide attention from the community. However, in medical domain, radiology report generatio…
cs.CV2024
From Recognition to Prediction: Leveraging Sequence Reasoning for Action Anticipation
Xin Liu, Chao Hao, Zitong Yu +2
The action anticipation task refers to predicting what action will happen based on observed videos, which requires the model to have a strong ability to summarize the present and t…
cs.CV2024★ 3 cited
A Simple yet Effective Network based on Vision Transformer for Camouflaged Object and Salient Object Detection
Chao Hao, Zitong Yu, Xin Liu +3
Camouflaged object detection (COD) and salient object detection (SOD) are two distinct yet closely-related computer vision tasks widely studied during the past decades. Though shar…