7 papers
GET: Generative Embedding Translation for Medical Image Segmentation
Md Maklachur Rahman, Md Hasan Al Banna, Saraf Anjum +2
Generative segmentation provides an alternative to direct pixel-wise prediction by operating on learned latent representations, but effective image-to-mask translation must preserv…
Dual-Domain Cross-Modal Decoding for Clinical Text-Guided Medical Image Segmentation
Md Maklachur Rahman, Tracy Hammond
Clinical text can narrow down what to segment, but recent text-guided designs emphasize spatial alignment while overlooking frequency content that governs texture and boundaries. W…
MambaLiteUNet: Cross-Gated Adaptive Feature Fusion for Robust Skin Lesion Segmentation
Md Maklachur Rahman, Soon Ki Jung, Tracy Hammond
Recent segmentation models have demonstrated promising efficiency by aggressively reducing parameter counts and computational complexity. However, these models often struggle to ac…
Masked Contrastive Pre-Training Improves Music Audio Key Detection
Ori Yonay, Tracy Hammond, Tianbao Yang
Self-supervised music foundation models underperform on key detection, which requires pitch-sensitive representations. In this work, we present the first systematic study showing t…
DependencyAI: Detecting AI Generated Text through Dependency Parsing
Sara Ahmed, Tracy Hammond
As large language models (LLMs) become increasingly prevalent, reliable methods for detecting AI-generated text are critical for mitigating potential risks. We introduce Dependency…
Myna: Masking-Based Contrastive Learning of Musical Representations
Ori Yonay, Tracy Hammond, Tianbao Yang
We present Myna, a simple yet effective approach for self-supervised musical representation learning. Built on a contrastive learning framework, Myna introduces two key innovations…