From the 2 of 14 linked papers with an AI index.
14 papers
Rad-JEPA 3D: Radiology Joint-Embedding Predictive Model for 3D Computed Tomography
Quoc-Huy Trinh, Minh-Van Nguyen, Ulas Bagci
The paper presents Rad-JEPA 3D, a self‑supervised joint‑embedding model that learns 3D CT representations by predicting latent features of a full scan from a masked view, using a h…
How Context Attribution Handles What the Model Already Knows
Quoc-Huy Trinh, Lin Zhu, Sebastian Szyller
Context attribution methods for large language models (LLMs) identify which input context contributes to the model response. Recent works show the initial success in attributing th…
Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space
Quoc-Huy Trinh, Xi Ding, Yang Liu +7
The paper introduces SpatialMed, a benchmark and an automated pipeline that generates 3D spatial visual question‑answer pairs for medical imaging, and shows that current multimodal…
SRMA-Mamba: Spatial Reverse Mamba Attention Network for Pathological Liver Segmentation in MRI Volumes
Jun Zeng, Quoc-Huy Trinh, Deepak Ranjan Nayak +3
Liver cirrhosis plays a critical role in the prognosis of chronic liver disease. Early detection and timely intervention are essential for reducing mortality rates. However, the in…
Firebolt-VL: Efficient Vision-Language Understanding with Cross-Modality Modulation
Quoc-Huy Trinh, Mustapha Abdullahi, Bo Zhao +1
Recent advances in multimodal large language models (MLLMs) have enabled impressive progress in vision-language understanding, yet their high computational cost limits deployment i…
PRS-Med: Position Reasoning Segmentation in Medical Imaging
Quoc-Huy Trinh, Minh-Van Nguyen, Jun Zeng +2
Prompt-based medical image segmentation has rapidly emerged, yet existing methods rely on explicit prompts like bounding boxes and struggle to reason about the spatial relationship…