Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
ViDRiP-LLaVA: A Dataset and Benchmark for Diagnostic Reasoning from Pathology Videos
Trinh T. L. Vuong, Jin Tae Kwak
We present ViDRiP-LLaVA, the first large multimodal model (LMM) in computational pathology that integrates three distinct image scenarios, including single patch images, automatica…
cs.CV2024
QuIIL at T3 challenge: Towards Automation in Life-Saving Intervention Procedures from First-Person View
Trinh T. L. Vuong, Doanh C. Bui, Jin Tae Kwak
In this paper, we present our solutions for a spectrum of automation tasks in life-saving intervention procedures within the Trauma THOMPSON (T3) Challenge, encompassing action rec…
cs.CV2024
Towards a text-based quantitative and explainable histopathology image analysis
Anh Tien Nguyen, Trinh Thi Le Vuong, Jin Tae Kwak
Recently, vision-language pre-trained models have emerged in computational pathology. Previous works generally focused on the alignment of image-text pairs via the contrastive pre-…