works on

From the 2 of 14 linked papers with an AI index.

activity
20242026
collaborators

14 papers

cs.CV2026

Rad-JEPA 3D: Radiology Joint-Embedding Predictive Model for 3D Computed Tomography

Quoc-Huy Trinh, Minh-Van Nguyen, Ulas Bagci

The paper presents Rad-JEPA 3D, a self‑supervised joint‑embedding model that learns 3D CT representations by predicting latent features of a full scan from a masked view, using a h…

cs.CL2026

How Context Attribution Handles What the Model Already Knows

Quoc-Huy Trinh, Lin Zhu, Sebastian Szyller

Context attribution methods for large language models (LLMs) identify which input context contributes to the model response. Recent works show the initial success in attributing th…

cs.CV2026

Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space

Quoc-Huy Trinh, Xi Ding, Yang Liu +7

The paper introduces SpatialMed, a benchmark and an automated pipeline that generates 3D spatial visual question‑answer pairs for medical imaging, and shows that current multimodal…

cs.CV2026

SRMA-Mamba: Spatial Reverse Mamba Attention Network for Pathological Liver Segmentation in MRI Volumes

Jun Zeng, Quoc-Huy Trinh, Deepak Ranjan Nayak +3

Liver cirrhosis plays a critical role in the prognosis of chronic liver disease. Early detection and timely intervention are essential for reducing mortality rates. However, the in…

cs.CV2026

Firebolt-VL: Efficient Vision-Language Understanding with Cross-Modality Modulation

Quoc-Huy Trinh, Mustapha Abdullahi, Bo Zhao +1

Recent advances in multimodal large language models (MLLMs) have enabled impressive progress in vision-language understanding, yet their high computational cost limits deployment i…

cs.CV2026

PRS-Med: Position Reasoning Segmentation in Medical Imaging

Quoc-Huy Trinh, Minh-Van Nguyen, Jun Zeng +2

Prompt-based medical image segmentation has rapidly emerged, yet existing methods rely on explicit prompts like bounding boxes and struggle to reason about the spatial relationship…