5 papers
Agentic Visual Reasoning in Whole-Slide Pathology Images via Active Perception
Jingyun Chen, Fengchun Liu, Linghan Cai +5
Whole-slide visual reasoning requires identifying sparse diagnostic evidence in gigapixel pathology slides and integrating observations across spatial scales. Existing WSI methods…
PathFLIP: Fine-grained Language-Image Pretraining for Versatile Computational Pathology
Fengchun Liu, Songhan Jiang, Linghan Cai +2
While Vision-Language Models (VLMs) have achieved notable progress in computational pathology (CPath), the gigapixel scale and spatial heterogeneity of Whole Slide Images (WSIs) co…
MedAction: Towards Active Multi-turn Clinical Diagnostic LLMs
Hsin-Ling Hsu, Zizheng Wang, Donghua Zhang +9
Most existing LLM diagnoses are evaluated on static, single-turn settings where complete patient information is provided upfront, an oversimplification of real clinical practice. W…
DINO-MVR: Multi-View Readout of Frozen DINOv3 for Annotation-Efficient Medical Segmentation
Wei Jiang, Feng Liu, Nan Ye +1
Adapting foundation models to medical segmentation typically requires either backbone fine-tuning or high-capacity task-specific decoders, both of which are difficult to fit reliab…
PathReasoner-R1: Instilling Structured Reasoning into Pathology Vision-Language Model via Knowledge-Guided Policy Optimization
Songhan Jiang, Fengchun Liu, Ziyue Wang +2
Vision-Language Models (VLMs) are advancing computational pathology with superior visual understanding capabilities. However, current systems often reduce diagnosis to directly out…