3 papers
cs.CV2026
Rethinking MLLM Itself as a Segmenter with a Single Segmentation Token
Anqi Zhang, Xiaokang Ji, Guangyu Gao +3
Recent segmentation methods leveraging Multi-modal Large Language Models (MLLMs) have shown reliable object-level segmentation and enhanced spatial perception. However, almost all…
cs.CV2024
Few Exemplar-Based General Medical Image Segmentation via Domain-Aware Selective Adaptation
Chen Xu, Qiming Huang, Yuqi Hou +4
Medical image segmentation poses challenges due to domain gaps, data modality variations, and dependency on domain knowledge or experts, especially for low- and middle-income count…
cs.CV2024
Frame Interpolation with Consecutive Brownian Bridge Diffusion
Zonglin Lyu, Ming Li, Jianbo Jiao +1
Recent work in Video Frame Interpolation (VFI) tries to formulate VFI as a diffusion-based conditional image generation problem, synthesizing the intermediate frame given a random…