-
Detecting gamma-band responses to the speech envelope for the ICASSP 2024 Auditory EEG Decoding Signal Processing Grand Challenge
Authors:
Mike Thornton,
Jonas Auernheimer,
Constantin Jehn,
Danilo Mandic,
Tobias Reichenbach
Abstract:
The 2024 ICASSP Auditory EEG Signal Processing Grand Challenge concerns the decoding of electroencephalography (EEG) measurements taken from participants who listened to speech material. This work details our solution to the match-mismatch sub-task: given a short temporal segment of EEG recordings and several candidate speech segments, the task is to classify which of the speech segments was time-…
▽ More
The 2024 ICASSP Auditory EEG Signal Processing Grand Challenge concerns the decoding of electroencephalography (EEG) measurements taken from participants who listened to speech material. This work details our solution to the match-mismatch sub-task: given a short temporal segment of EEG recordings and several candidate speech segments, the task is to classify which of the speech segments was time-aligned with the EEG signals. We show that high-frequency gamma-band responses to the speech envelope can be detected with a high accuracy. By jointly assessing gamma-band responses and low-frequency envelope tracking, we develop a match-mismatch decoder which placed first in this task.
△ Less
Submitted 30 January, 2024;
originally announced January 2024.
-
Comparison of linear and nonlinear methods for decoding selective attention to speech from ear-EEG recordings
Authors:
Mike Thornton,
Danilo Mandic,
Tobias Reichenbach
Abstract:
Many people with hearing loss struggle to comprehend speech in crowded auditory scenes, even when they are using hearing aids. It has recently been demonstrated that the focus of a listener's selective attention to speech can be decoded from their electroencephalography (EEG) recordings, raising the prospect of smart EEG-steered hearing aids which restore speech comprehension in adverse acoustic e…
▽ More
Many people with hearing loss struggle to comprehend speech in crowded auditory scenes, even when they are using hearing aids. It has recently been demonstrated that the focus of a listener's selective attention to speech can be decoded from their electroencephalography (EEG) recordings, raising the prospect of smart EEG-steered hearing aids which restore speech comprehension in adverse acoustic environments (such as the cocktail party). To this end, we here assess the feasibility of using a novel, ultra-wearable ear-EEG device to classify the selective attention of normal-hearing listeners who participated in a two-talker competing-speakers experiment.
Eighteen participants took part in a diotic listening task, whereby they were asked to attend to one narrator whilst ignoring the other. Encoding models were estimated from the recorded signals, and these confirmed that the device has the ability to capture auditory responses that are consistent with those reported in high-density EEG studies. Several state-of-the-art auditory attention decoding algorithms were next compared, including stimulus-reconstruction algorithms based on linear regression as well as non-linear deep neural networks, and canonical correlation analysis (CCA).
Meaningful markers of selective auditory attention could be extracted from the ear-EEG signals of all 18 participants, even when those markers were derived from relatively short EEG segments of just five seconds in duration. Algorithms which related the EEG signals to the rising edges of the speech temporal envelope (onset envelope) were more successful than those which made use of the temporal envelope itself. The CCA algorithm achieved the highest mean attention decoding accuracy, although differences between the performances of the three algorithms were both small and not statistically significant when EEG segments of short durations were employed.
△ Less
Submitted 15 November, 2024; v1 submitted 10 January, 2024;
originally announced January 2024.
-
Decoding Envelope and Frequency-Following EEG Responses to Continuous Speech Using Deep Neural Networks
Authors:
Mike Thornton,
Danilo Mandic,
Tobias Reichenbach
Abstract:
The electroencephalogram (EEG) offers a non-invasive means by which a listener's auditory system may be monitored during continuous speech perception. Reliable auditory-EEG decoders could facilitate the objective diagnosis of hearing disorders, or find applications in cognitively-steered hearing aids. Previously, we developed decoders for the ICASSP Auditory EEG Signal Processing Grand Challenge (…
▽ More
The electroencephalogram (EEG) offers a non-invasive means by which a listener's auditory system may be monitored during continuous speech perception. Reliable auditory-EEG decoders could facilitate the objective diagnosis of hearing disorders, or find applications in cognitively-steered hearing aids. Previously, we developed decoders for the ICASSP Auditory EEG Signal Processing Grand Challenge (SPGC). These decoders aimed to solve the match-mismatch task: given a short temporal segment of EEG recordings, and two candidate speech segments, the task is to identify which of the two speech segments is temporally aligned, or matched, with the EEG segment. The decoders made use of cortical responses to the speech envelope, as well as speech-related frequency-following responses, to relate the EEG recordings to the speech stimuli. Here we comprehensively document the methods by which the decoders were developed. We extend our previous analysis by exploring the association between speaker characteristics (pitch and sex) and classification accuracy, and provide a full statistical analysis of the final performance of the decoders as evaluated on a heldout portion of the dataset. Finally, the generalisation capabilities of the decoders are characterised, by evaluating them using an entirely different dataset which contains EEG recorded under a variety of speech-listening conditions. The results show that the match-mismatch decoders achieve accurate and robust classification accuracies, and they can even serve as auditory attention decoders without additional training.
△ Less
Submitted 15 December, 2023;
originally announced December 2023.
-
Relating EEG recordings to speech using envelope tracking and the speech-FFR
Authors:
Mike Thornton,
Danilo Mandic,
Tobias Reichenbach
Abstract:
During speech perception, a listener's electroencephalogram (EEG) reflects acoustic-level processing as well as higher-level cognitive factors such as speech comprehension and attention. However, decoding speech from EEG recordings is challenging due to the low signal-to-noise ratios of EEG signals. We report on an approach developed for the ICASSP 2023 'Auditory EEG Decoding' Signal Processing Gr…
▽ More
During speech perception, a listener's electroencephalogram (EEG) reflects acoustic-level processing as well as higher-level cognitive factors such as speech comprehension and attention. However, decoding speech from EEG recordings is challenging due to the low signal-to-noise ratios of EEG signals. We report on an approach developed for the ICASSP 2023 'Auditory EEG Decoding' Signal Processing Grand Challenge. A simple ensembling method is shown to considerably improve upon the baseline decoder performance. Even higher classification rates are achieved by jointly decoding the speech-evoked frequency-following response and responses to the temporal envelope of speech, as well as by fine-tuning the decoders to individual subjects. Our results could have applications in the diagnosis of hearing disorders or in cognitively steered hearing aids.
△ Less
Submitted 11 March, 2023;
originally announced March 2023.