9 papers
Structured Phonological Representations for Audio-Articulatory rtMRI Speech Classification
Abner Hernandez, Tomás Arias Vergara, Daiqi Liu +2
Real-time MRI makes it possible to observe vocal-tract articulation during speech, but mapping these articulatory patterns to phonetic and phonological categories remains challengi…
WING: A Window-Prior-Based Generative Network with Gated Inception for Cross-Modality CT Synthesis
Siyuan Mei, Yan Xia, Yipeng Sun +7
Generating CT volumes from MRI and CBCT can improve treatment planning in adaptive radiotherapy while avoiding additional radiation exposure. However, direct regression of CT inten…
Multilingual Phonological Feature Recognition with Self-Supervised Speech Models
Abner Hernandez, Tomás Arias-Vergara, Daiqi Liu +2
Phonological features provide a language-general and linguistically grounded representation of speech. We present PhonoQ-2.0, a multilingual frame-level phonological feature recogn…
Speech-Guided Multimodal Learning for Vocal Tract Segmentation in Real-Time MRI
Daiqi Liu, Lukas Mulzer, Md Hasan +11
Segmenting vocal tract articulators in real-time MRI (rtMRI) is a challenging dynamic image segmentation problem characterized by low contrast, rapid motion, and limited spatial re…
SIREM: Speech-Informed MRI Reconstruction with Learned Sampling
Md Hasan, Nyvenn Castro, Daiqi Liu +6
Real-time magnetic resonance imaging (rtMRI) of speech production enables non-invasive visualization of dynamic vocal-tract motion and is valuable for speech science and clinical a…
VocSegMRI: Multimodal Learning for Precise Vocal Tract Segmentation in Real-time MRI
Daiqi Liu, Johannes Enk, Maureen Stone +7
Accurate segmentation of articulatory structures in real-time MRI (rtMRI) remains challenging, as existing methods rely primarily on visual cues and overlook complementary informat…