activity
20222024
most citedVideoLLM: Modeling Video Sequence with Large Language Models

15 citations · 21 across the 8 of their papers we have counts for

collaborators

8 papers

eess.IV2024

Advancing COVID-19 Detection in 3D CT Scans

Qingqiu Li, Runtian Yuan, Junlin Hou +4

To make a more accurate diagnosis of COVID-19, we propose a straightforward yet effective model. Firstly, we analyse the characteristics of 3D CT scans and remove the non-lung part…

eess.IV2024

Domain Adaptation Using Pseudo Labels for COVID-19 Detection

Runtian Yuan, Qingqiu Li, Junlin Hou +4

In response to the need for rapid and accurate COVID-19 diagnosis during the global pandemic, we present a two-stage framework that leverages pseudo labels for domain adaptation to…

cs.CV20241 cited

Anatomical Structure-Guided Medical Vision-Language Pre-training

Qingqiu Li, Xiaohan Yan, Jilan Xu +6

Learning medical visual representations through vision-language pre-training has reached remarkable progress. Despite the promising performance, it still faces challenges, i.e., lo…

cs.CV2023

Enhanced Knowledge Injection for Radiology Report Generation

Qingqiu Li, Jilan Xu, Runtian Yuan +5

Automatic generation of radiology reports holds crucial clinical value, as it can alleviate substantial workload on radiologists and remind less experienced ones of potential anoma…

cs.CV202315 cited

VideoLLM: Modeling Video Sequence with Large Language Models

Guo Chen, Yin-Dong Zheng, Jiahao Wang +8

With the exponential growth of video data, there is an urgent need for automated technology to analyze and comprehend video content. However, existing video understanding models ar…

cs.CV2023

Mask Hierarchical Features For Self-Supervised Learning

Fenggang Liu, Yangguang Li, Feng Liang +3

This paper shows that Masking the Deep hierarchical features is an efficient self-supervised method, denoted as MaskDeep. MaskDeep treats each patch in the representation space as…