Abstract:Objective. Smart hearing aids which can decode the focus of a user's attention could considerably improve comprehension levels in noisy environments. Methods for decoding auditory attention from electroencephalography (EEG) have attracted considerable interest for this reason. Recent studies suggest that the integration of deep neural networks (DNNs) into existing auditory attention decoding algorithms is highly beneficial, although it remains unclear whether these enhanced algorithms can perform robustly in different real-world scenarios. To this end, we sought to characterise the performance of DNNs at reconstructing the envelope of an attended speech stream from EEG recordings in different listening conditions. In addition, given the relatively sparse availability of EEG data, we investigate possibility of applying subject-independent algorithms to EEG recorded from unseen individuals. Approach. Both linear models and nonlinear DNNs were employed to decode the envelope of clean speech from EEG recordings, with and without subject-specific information. The mean behaviour, as well as the variability of the reconstruction, was characterised for each model. We then trained subject-specific linear models and DNNs to reconstruct the envelope of speech in clean and noisy conditions, and investigated how well they performed in different listening scenarios. We also established that these models can be used to decode auditory attention in competing-speaker scenarios. Main results. The DNNs offered a considerable advantage over their linear counterpart at reconstructing the envelope of clean speech. This advantage persisted even when subject-specific information was unavailable at the time of training. The same DNN architectures generalised to a distinct dataset, which contained EEG recorded under a variety of listening conditions. In competing-speakers and speech-in-noise conditions, the DNNs significantly outperformed the linear models. Finally, the DNNs offered a considerable improvement over the linear approach at decoding auditory attention in competing-speakers scenarios. Significance. We present the first detailed study into the extent to which DNNs can be employed for reconstructing the envelope of an attended speech stream. We conclusively demonstrate that DNNs have the ability to improve the reconstruction of the attended speech envelope. The variance of the reconstruction error is shown to be similar for both DNNs and the linear model. Overall, DNNs are demonstrated to show promise for real-world auditory attention decoding, since they perform well in multiple listening conditions and generalise to data recorded from unseen participants.

Comparison of linear and nonlinear methods for decoding selective attention to speech from ear-EEG recordings

Decoding auditory attention (in real time) with eeg

Comparison of Two-Talker Attention Decoding from EEG with Nonlinear Neural Networks and Linear Methods

Decoding Selective Attention in Normal Hearing Listeners and Bilateral Cochlear Implant Users With Concealed Ear EEG

EEG-based Auditory Attention Decoding: Towards Neuro-Steered Hearing Devices

Auditory attention decoding from electroencephalography based on long short-term memory networks

EEG decoding of the target speaker in a cocktail party scenario: considerations regarding dynamic switching of talker location

'Are you even listening?' - EEG-based decoding of absolute auditory attention to natural speech

Linear versus deep learning methods for noisy speech separation for EEG-informed attention decoding

Decoding of selective attention to continuous speech from the human auditory brainstem response

Predicting speech intelligibility from a selective attention decoding paradigm in cochlear implant users

Deep learning-based auditory attention decoding in listeners with hearing impairment

Neural decoding of attentional selection in multi-speaker environments without access to clean sources

Using Ear-EEG to Decode Auditory Attention in Multiple-speaker Environment

Fast EEG-Based Decoding Of The Directional Focus Of Auditory Attention Using Common Spatial Patterns

Real-time control of a hearing instrument with EEG-based attention decoding

Robust decoding of the speech envelope from EEG recordings through deep neural networks

Towards Estimating Selective Auditory Attention from EEG Using a Novel Time-Frequency-synchronisation Framework

Toward Decoding Selective Attention From Single-Trial EEG Data in Cochlear Implant Users

Ear-EEG Measures of Auditory Attention to Continuous Speech