collaborators

9 papers

cs.CL2026

Structured Phonological Representations for Audio-Articulatory rtMRI Speech Classification

Abner Hernandez, Tomás Arias Vergara, Daiqi Liu +2

Real-time MRI makes it possible to observe vocal-tract articulation during speech, but mapping these articulatory patterns to phonetic and phonological categories remains challengi…

cs.CV2026

WING: A Window-Prior-Based Generative Network with Gated Inception for Cross-Modality CT Synthesis

Siyuan Mei, Yan Xia, Yipeng Sun +7

Generating CT volumes from MRI and CBCT can improve treatment planning in adaptive radiotherapy while avoiding additional radiation exposure. However, direct regression of CT inten…

cs.CL2026

Multilingual Phonological Feature Recognition with Self-Supervised Speech Models

Abner Hernandez, Tomás Arias-Vergara, Daiqi Liu +2

Phonological features provide a language-general and linguistically grounded representation of speech. We present PhonoQ-2.0, a multilingual frame-level phonological feature recogn…

cs.CV2026

Speech-Guided Multimodal Learning for Vocal Tract Segmentation in Real-Time MRI

Daiqi Liu, Lukas Mulzer, Md Hasan +11

Segmenting vocal tract articulators in real-time MRI (rtMRI) is a challenging dynamic image segmentation problem characterized by low contrast, rapid motion, and limited spatial re…

cs.SD2026

SIREM: Speech-Informed MRI Reconstruction with Learned Sampling

Md Hasan, Nyvenn Castro, Daiqi Liu +6

Real-time magnetic resonance imaging (rtMRI) of speech production enables non-invasive visualization of dynamic vocal-tract motion and is valuable for speech science and clinical a…

cs.CV2026

VocSegMRI: Multimodal Learning for Precise Vocal Tract Segmentation in Real-time MRI

Daiqi Liu, Johannes Enk, Maureen Stone +7

Accurate segmentation of articulatory structures in real-time MRI (rtMRI) remains challenging, as existing methods rely primarily on visual cues and overlook complementary informat…