activity
20222024
most citedTowards Medical Artificial General Intelligence via Knowledge-Enhanced Multimodal Pretraining

9 citations · 12 across the 5 of their papers we have counts for

collaborators

5 papers

cs.CV2024

VidMan: Exploiting Implicit Dynamics from Video Diffusion Model for Effective Robot Manipulation

Youpeng Wen, Junfan Lin, Yi Zhu +4

Recent advancements utilizing large-scale video data for learning video generation models demonstrate significant potential in understanding complex physical dynamics. It suggests…

cs.CV2023

Learning Snippet-to-Motion Progression for Skeleton-based Human Motion Prediction

Xinshun Wang, Qiongjie Cui, Chen Chen +2

Existing Graph Convolutional Networks to achieve human motion prediction largely adopt a one-step scheme, which output the prediction straight from history input, failing to exploi…

cs.AI20239 cited

Towards Medical Artificial General Intelligence via Knowledge-Enhanced Multimodal Pretraining

Bingqian Lin, Zicong Chen, Mingjie Li +13

Medical artificial general intelligence (MAGI) enables one foundation model to solve different medical tasks, which is very practical in the medical domain. It can significantly re…

cs.CV20232 cited

CapDet: Unifying Dense Captioning and Open-World Detection Pretraining

Yanxin Long, Youpeng Wen, Jianhua Han +5

Benefiting from large-scale vision-language pre-training on image-text pairs, open-world detection methods have shown superior generalization ability under the zero-shot or few-sho…

cs.CV20221 cited

PCCT: Progressive Class-Center Triplet Loss for Imbalanced Medical Image Classification

Kanghao Chen, Weixian Lei, Rong Zhang +3

Imbalanced training data is a significant challenge for medical image classification. In this study, we propose a novel Progressive Class-Center Triplet (PCCT) framework to allevia…