2 papers
cs.CV2025
Unlocking the Forgery Detection Potential of Vanilla MLLMs: A Novel Training-Free Pipeline
Rui Zuo, Qinyue Tong, Zhe-Ming Lu +1
With the rapid advancement of artificial intelligence-generated content (AIGC) technologies, including multimodal large language models (MLLMs) and diffusion models, image generati…
cs.CV2025
MediRound: Multi-Round Entity-Level Reasoning Segmentation in Medical Images
Qinyue Tong, Ziqian Lu, Jun Liu +3
Despite notable progress in text-guided medical image segmentation nowadays, these methods are limited to single-round dialogues and fail to support multi-round reasoning, which is…