Nowadays, the number of people with hearing loss is increasing rapidly and is estimated to reach around 2.5 billion worldwide by 2050, according to a report from the World Health Organization. Hearing loss severely affects both the physical and mental...
Nowadays, the number of people with hearing loss is increasing rapidly and is estimated to reach around 2.5 billion worldwide by 2050, according to a report from the World Health Organization. Hearing loss severely affects both the physical and mental health of patients. For instance, it has a negative impact on cognitive and language development in infants and children, and it can significantly harm adults’ mental health by causing social isolation and unemployment. To address this issue, one of the current solutions is to provide patients with hearing assistive devices. These devices function by amplifying speech sounds while suppressing background noise. However, in multi-talker situations (commonly referred to as the cocktail party scenario), these devices struggle to detect and amplify the attended speech. To overcome this challenge, Auditory Attention Decoding (AAD) has been introduced, as it can infer the attended speaker by decoding brain signals. Although AAD has been widely studied and developed under laboratory conditions, several limitations still question its feasibility in real-world applications. First, under passive listening conditions, when the listener is not consciously attending to the incoming sound, can the decoding algorithm still operate effectively? Second, most existing AAD algorithms are trained in a supervised manner and remain fixed during operation. This setup may limit their performance in real-world environments, where both the acoustic surroundings and brain signals are highly dynamic. This dissertation addresses these two challenges. First, it investigates the feasibility of using brain signals to decode auditory attention by analyzing neural tracking of the speech envelope under extreme passive listening conditions. Second, it proposes a fast, time-adaptive, and unsupervised AAD algorithm designed for plug-and-play operation. The results of this study demonstrate the feasibility of applying AAD in real-world environments, showing that it can function effectively under everyday listening conditions. Furthermore, the proposed algorithm performs well in a plug-and-play manner, highlighting its superiority and potential for integration into future hearing technologies.