cs.SDOct 7, 2026

A multi-scenario EEG dataset for auditory attention decoding in naturalistic multi-talker environments

Authors: Shu Peng, Rui Liu, Yufei Zhang, Wenlong You, Zhige Chen, Jiachen Xi, Qiyuan Sun, Yan Liu, +2 more

Organizations: Department of Data Science and Artificial Intelligence, The Hong Kong Polytechnic University, Hong Kong SAR, China · Department of Computing, The Hong Kong Polytechnic University, Hong Kong SAR, China

Abstract

Understanding how the brain selectively follows relevant speech amid competing voices is a central challenge in auditory neuroscience and a key step toward neuro-steered hearing technologies. However, most open-source Electroencephalography (EEG) datasets for Auditory Attention Decoding (AAD) use idealized single-competing-talker paradigms that oversimplify the acoustic, spatial, and semantic structure of everyday communication. To capture this ecological complexity, we introduce the SoundBubble-EEG dataset: a high-density 128-channel EEG resource comprising more than 25 hours of recordings from 30 participants. The paradigm requires listeners to selectively attend to a dynamic target speaker group, a designated "sound bubble", amid competing multi-speaker distractor bubbles across three realistic scenarios: a restaurant, a home TV viewing, and a meeting discussion. By bridging the gap between constrained laboratory protocols and real-world auditory scenes, this dataset enables investigations of multi-talker speech comprehension, neural speech tracking, and cross-scenario generalization. It also provides a benchmark for AAD algorithms under realistic acoustic and semantic variability and may support auditory neuroscience and the development of neuro-steered hearing technologies.

Figures & tables

Explore similar work

CardsList
  1. FAConformer: Frequency-Aware Convolutional Transformer for Auditory Attention Decoding

    Jun 12, 2026Ziwei Wang, Xingyi He, Tianwang Jia +2Auditory AttentionNeural Audio

  2. NEUROTOKEN: Joint Source and Directional AAD with Envelope Decoding via Conditional Flow Matching

    Sep 30, 2026Ali Alavi, Donald S. WilliamsonAuditory AttentionJoint Search

  3. Decoding Stimulus Reconstruction-Based Auditory Attention Robustly in Unbalanced EEG Datasets

    May 25, 2026Yuanming Zhang, Yayun Liang, Zhibin Lin +1Auditory AttentionSmall-Sample Electroencephalogram Datasets