FastText and GPT-2 embeddings linearly predict sEEG high-gamma responses during word reading with high correlation, but Wav2Vec 2.0 performs poorly in two participants.
NeuroIncept Decoder for High-Fidelity Speech Reconstruction from Neural Activity
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
This paper introduces a novel algorithm designed for speech synthesis from neural activity recordings obtained using invasive electroencephalography (EEG) techniques. The proposed system offers a promising communication solution for individuals with severe speech impairments. Central to our approach is the integration of time-frequency features in the high-gamma band computed from EEG recordings with an advanced NeuroIncept Decoder architecture. This neural network architecture combines Convolutional Neural Networks (CNNs) and Gated Recurrent Units (GRUs) to reconstruct audio spectrograms from neural patterns. Our model demonstrates robust mean correlation coefficients between predicted and actual spectrograms, though inter-subject variability indicates distinct neural processing mechanisms among participants. Overall, our study highlights the potential of neural decoding techniques to restore communicative abilities in individuals with speech disorders and paves the way for future advancements in brain-computer interface technologies.
citation-role summary
citation-polarity summary
fields
cs.HC 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Recreating Neural Activity During Speech Production with Language and Speech Model Embeddings
FastText and GPT-2 embeddings linearly predict sEEG high-gamma responses during word reading with high correlation, but Wav2Vec 2.0 performs poorly in two participants.