411 citations · 558 across the 6 of their papers we have counts for
5 papers · 1 filter
Biomedical SAM 2: Segment Anything in Biomedical Images and Videos
Zhiling Yan, Weixiang Sun, Rong Zhou +8
Medical image segmentation and video object segmentation are essential for diagnosing and analyzing diseases by identifying and measuring biological structures. Recent advances in…
Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models
Yixin Liu, Kai Zhang, Yuan Li +9
Sora is a text-to-video generative AI model, released by OpenAI in February 2024. The model is trained to generate videos of realistic or imaginative scenes from text instructions…
Multimodal ChatGPT for Medical Applications: an Experimental Study of GPT-4V
Zhiling Yan, Kai Zhang, Rong Zhou +3
In this paper, we critically evaluate the capabilities of the state-of-the-art multimodal large language model, i.e., GPT-4 with Vision (GPT-4V), on Visual Question Answering (VQA)…
MA-SAM: Modality-agnostic SAM Adaptation for 3D Medical Image Segmentation
Cheng Chen, Juzheng Miao, Dufan Wu +10
The Segment Anything Model (SAM), a foundation model for general image segmentation, has demonstrated impressive zero-shot performance across numerous natural image segmentation ta…
Learning to Generate Poetic Chinese Landscape Painting with Calligraphy
Shaozu Yuan, Aijun Dai, Zhiling Yan +5
In this paper, we present a novel system (denoted as Polaca) to generate poetic Chinese landscape painting with calligraphy. Unlike previous single image-to-image painting generati…