4 citations · 5 across the 4 of their papers we have counts for
4 papers
Visual Prompting in Multimodal Large Language Models: A Survey
Junda Wu, Zhehao Zhang, Yu Xia +12
Multimodal large language models (MLLMs) equip pre-trained large-language models (LLMs) with visual capabilities. While textual prompting in LLMs has been widely studied, visual pr…
Deep Deformable Models: Learning 3D Shape Abstractions with Part Consistency
Di Liu, Long Zhao, Qilong Zhangli +3
The task of shape abstraction with semantic part consistency is challenging due to the complex geometries of natural objects. Recent methods learn to represent an object shape usin…
Pathology-and-genomics Multimodal Transformer for Survival Outcome Prediction
Kexin Ding, Mu Zhou, Dimitris N. Metaxas +1
Survival outcome assessment is challenging and inherently associated with multiple clinical factors (e.g., imaging and genomics biomarkers) in cancer. Enabling multimodal analytics…
Contrastive and Selective Hidden Embeddings for Medical Image Segmentation
Zhuowei Li, Zihao Liu, Zhiqiang Hu +5
Medical image segmentation has been widely recognized as a pivot procedure for clinical diagnosis, analysis, and treatment planning. However, the laborious and expensive annotation…