From the 2 of 5 linked papers with an AI index.
5 papers
Phoneme- vs. Character-Level Targets and Selective State-Space Models for Intracortical Brain-to-Text
Lucas Zamora Vera, Jose A. Gonzalez-Lopez
The paper compares phoneme- versus character-level output targets and evaluates selective state-space (Mamba) versus recurrent (GRU) decoders for intracortical brain-to-text system…
Zero-Shot Face-to-Speech Synthesis via Latent Space Adaptation of a Style-Diffusion TTS Model
Carlos Muñoz-Romero, Jose A. Gonzalez-Lopez
The paper presents a zero-shot Face-to-Speech system that generates a plausible voice from a single facial image by adapting a frozen StyleTTS 2 model with a lightweight face adapt…
End-to-End Intracortical Speech Decoding from Neural Activity
Owais Mujtaba Khanday, Jose A. Gonzalez-Lopez, Marc Ouellet +2
Current high-performing intracortical speech neuroprostheses achieve low word error rates but typically rely on external language models during inference, increasing memory, comput…
Recreating Neural Activity During Speech Production with Language and Speech Model Embeddings
Owais Mujtaba Khanday, Pablo Rodroguez San Esteban, Zubair Ahmad Lone +2
Understanding how neural activity encodes speech and language production is a fundamental challenge in neuroscience and artificial intelligence. This study investigates whether emb…
NeuroIncept Decoder for High-Fidelity Speech Reconstruction from Neural Activity
Owais Mujtaba Khanday, José L. Pérez-Córdoba, Mohd Yaqub Mir +2
This paper introduces a novel algorithm designed for speech synthesis from neural activity recordings obtained using invasive electroencephalography (EEG) techniques. The proposed…