{"dataset":{"id":"60751","dataset_id":"on005408","name":"The effect of speech masking on the human subcortical response to continuous speech","description":"This dataset contains EEG recordings used to derive auditory brainstem responses (ABRs) to continuous, naturally uttered speech under varying levels of speech masking. Data were collected from 25 normal-hearing adult participants presented with click trains and 'peaky speech' stimuli from one to five simultaneously presented talkers at different signal-to-noise ratios. The study aims to characterize how masking affects subcortical neural encoding of speech in human listeners, building on prior methods for deriving speech-ABRs.","owner_user_id":15,"status":"active","github_repo":"nemarDatasets/on005408","concept_doi":"10.82901/nemar.on005408","latest_version_doi":"10.82901/nemar.on005408.v1.0.0","created_at":"2026-06-27 01:01:40","updated_at":"2026-08-19 01:42:43","zenodo_concept_id":null,"is_sandbox":0,"visibility":"public","ezid_status":"public","enrichment_json":"{\n  \"version\": \"2.0\",\n  \"pipeline_stage\": \"validated\",\n  \"title\": \"The effect of speech masking on the human subcortical response to continuous speech\",\n  \"description\": \"This dataset contains EEG recordings used to derive auditory brainstem responses (ABRs) to continuous, naturally uttered speech under varying levels of speech masking. Data were collected from 25 normal-hearing adult participants presented with click trains and 'peaky speech' stimuli from one to five simultaneously presented talkers at different signal-to-noise ratios. The study aims to characterize how masking affects subcortical neural encoding of speech in human listeners, building on prior methods for deriving speech-ABRs.\",\n  \"methods_description\": \"EEG was recorded from participants seated in a darkened sound-isolating booth who rested or watched silent captioned videos. Stimuli included randomized click trains (40 Hz average rate, 60 x 10s trials) and peaky speech from up to five male narrators at four SNR levels (clean, 0 dB, -3 dB, -6 dB), presented at 65 dB SPL via ER-2 insert earphones connected to an RME Babyface Pro sound card at 48 kHz sampling rate. Stimulus presentation was controlled using custom Python scripts with expyfun. Raw EEG data are provided in BrainVision format (.eeg, .vhdr, .vmrk), with stimulus regressor files provided in HDF5 format.\",\n  \"license\": \"CC0\",\n  \"dataset_type\": \"raw\",\n  \"authors\": {\n    \"Melissa J. Polonenko\": {\n      \"orcid\": \"0000-0003-1914-6117\"\n    },\n    \"Ross K. Maddox\": {\n      \"orcid\": \"0000-0003-2668-0238\"\n    }\n  },\n  \"keywords\": [\n    {\n      \"term\": \"EEG\"\n    },\n    {\n      \"term\": \"Auditory Brainstem Response\"\n    },\n    {\n      \"term\": \"Speech Perception\",\n      \"subject_scheme\": \"MeSH\",\n      \"value_uri\": \"http://id.nlm.nih.gov/mesh/D013067\"\n    },\n    {\n      \"term\": \"auditory masking\"\n    },\n    {\n      \"term\": \"continuous speech encoding\"\n    },\n    {\n      \"term\": \"subcortical auditory processing\"\n    }\n  ],\n  \"related_identifiers\": [\n    {\n      \"identifier\": \"10.1523/ENEURO.0561-24.2025\",\n      \"identifier_type\": \"DOI\",\n      \"relation_type\": \"IsDescribedBy\"\n    },\n    {\n      \"identifier\": \"https://github.com/nemarDatasets/on005408\",\n      \"identifier_type\": \"URL\",\n      \"relation_type\": \"IsDescribedBy\"\n    },\n    {\n      \"identifier\": \"10.18112/openneuro.ds005408.v1.0.1\",\n      \"identifier_type\": \"DOI\",\n      \"relation_type\": \"IsDerivedFrom\"\n    },\n    {\n      \"identifier\": \"https://nemar.org/dataset/on005408\",\n      \"identifier_type\": \"URL\",\n      \"relation_type\": \"IsDescribedBy\"\n    }\n  ],\n  \"funding_references\": [\n    {\n      \"funder_name\": \"NIDCD\",\n      \"award_number\": \"R01DC017962\"\n    }\n  ],\n  \"resource_type_general\": \"Dataset\",\n  \"resource_type_specific\": \"EEG Dataset\",\n  \"modalities\": [\n    \"eeg\"\n  ],\n  \"sizes\": [\n    \"40.6 GB (2128 files)\"\n  ],\n  \"formats\": [\n    \".eeg\",\n    \".hdf5\",\n    \".json\",\n    \".md\",\n    \".tsv\",\n    \".vhdr\",\n    \".vmrk\",\n    \".yml\"\n  ],\n  \"source_hash\": \"b1a00b55e840b4996bf45132c0ab79ebfa1d0a28408f7193691ad1b8fb0d6fdc\"\n}","last_activity_at":"2026-06-27 01:01:40","source":"openneuro","source_id":"ds005408","subject_count":25,"modalities":"eeg","age_min":19,"age_max":37,"file_size":40639245713,"total_files":2250,"tasks":"peakysnr","metadata_columns_error":null,"staleness_warn_stage":null,"staleness_admin_notified_at":null,"authors":"Melissa J. Polonenko, Ross K. Maddox","license":"CC0","readme":"[![DOI](https://img.shields.io/badge/DOI-10.82901%2Fnemar.on005408-blue)](https://doi.org/10.82901/nemar.on005408)\n\n\nREADME\n------\nDetails related to access to the data\n-------------------------------------\nPlease contact the following authors for further information:\n    Melissa Polonenko (email: mpolonen@umn.edu) [corresponding author]\n    Ross Maddox (email: rkmaddox@med.umich.edu)\n\nOverview\n--------\nThis is the \"peaky_snr\" dataset for the paper by\nPolonenko MJ & Maddox RK, with citation listed below.\n\neNeuro: Polonenko, M. J., & Maddox, R. K. (2025). The effect of speech masking on the human subcortical response to continuous speech. eNeuro 24 March 2025, 12 (4) ENEURO.0561-24.2025; https://doi.org/10.1523/ENEURO.0561-24.2025\n\nBioRxiv: https://www.biorxiv.org/content/10.1101/2024.12.10.627771v1\n\nAuditory brainstem responses (ABRs) were derived to continuous peaky speech\nfrom between one up to five simultaneously presented talkers and from clicks.\nData was collected from June to July 2021.\n\n\nGoal: To better understand masking’s effects on the subcortical neural encoding\nof naturally uttered speech in human listeners.\n\nTo do this we leveraged our recently developed method for determining the\nauditory brainstem response (ABR) to speech (Polonenko and Maddox, 2021).\nWhereas our previous work was aimed at encoding of single talkers, here we\ndetermined the ABR to speech in quiet as well as in the presence of varying\nnumbers of other talkers. \n\nThe details of the experiment can be found at Polonenko & Maddox (2024). \n\nStimuli:\n    1) randomized click trains at an average rate of 40 Hz, \n    60 x 10 s trials for a total of 10 minutes;\n    2) peaky speech for up to 5 male narrators. 30 minutes of each SNR\n    (clean, 0 dB, -3 dB, -6 dB), corresponding to 1, 2, 3, and 5 talkers\n    presented simultaneously, each set to 65 dB. \n    \n    NOTE: files for each story were completely randomized. Random combinations\n    were created so that each story was equally represented in the data.\n\nThe code for stimulus preprocessing and EEG analysis is available on Github:\n    https://github.com/polonenkolab/peaky_snr\n\nFormat\n------\nThe dataset is formatted according to the EEG Brain Imaging Data Structure. It \nincludes EEG recording from participant 01 to 25 in raw brainvision format\n(3 files: .eeg, .vhdr, .vmrk) and stimuli files in format of .hdf5. The stimuli\nfiles contain the audio ('audio'), and regressors for the deconvolution \n('pinds' are the pulse indices, 'anm' is an auditory nerve model regressor,\n which was used during analyses but was not included as part of the article). \n\nGenerally, you can find detailed event data in the .tsv files and descriptions\nin the accompanying .json files. Raw eeg files are provided in the Brain\nProducts format.\n\nParticipants\n------------\n25 participants, mean ± SD age of 23.4 ± 5.5 years (19-37 years)\n\nInclusion criteria:\n    1) Age between 18-40 years\n    2) Normal hearing: audiometric thresholds 20 dB HL or better from 500 to 8000 Hz\n    3) Speak English as their primary language\n\nPlease see participants.tsv for more information.\n\nApparatus\n---------\nParticipants sat in a darkened sound-isolating booth and rested or watched\nsilent videos with closed captioning. Stimuli were presented at an average level\nof 65 dB SPL (per story; total for 5 talkers = 71 dB) and a sampling rate of \n48 kHz through ER-2 insert earphones plugged into an RME Babyface Pro digital\nsound card. Custom python scripts using expyfun were used to control the \nexperiment and stimulus presentation.\n\nDetails about the experiment\n----------------------------\nFor a detailed description of the task, see Polonenko & Maddox (2024) and the\nsupplied `task-peaky_snr_eeg.json` file. The 4 SNR speech conditions and the\nstory tokens were randomized. This means that the participant would not be able\nto follow the stories. For clicks the trials were not randomized\n(already random clicks).\n\nTrigger onset times in the tsv files have already been corrected for the tubing\ndelay of the insert earphones (but not in the events of the raw files).\nTriggers with values of \"1\" were recorded to the onset of the 10 s audio, and\nshortly after triggers with values of \"4\" or \"8\" were stamped to indicate info\nabout the trial. This was done by converting the decimal trial number to bits,\ndenoted b, then calculating 2 ** (b + 2). We've specified these trial triggers\nand more metadata of the events in each of the '*_eeg_events.tsv\" file, which\nis sufficient to know which trial corresponded to which type of stimulus\n(clicks or speech), snr, and which files of which stories were presented.\ne.g., alice_000_peaky_diotic_regress.hdf5 for the first file of the story\ncalled 'alice' (Alice in Wonderland).\n\n\n","bids_version":"1.7.0","sessions_count":null,"publish_date":null,"embedding_dirty":0,"license_tier":"public","zarr_status":"ready","zarr_converted_at":"2026-08-16 16:20:01","zarr_store_count":29,"zarr_index_etag":"208c3a785fde2696b7b19dfbcbc22321","zarr_source_commit":"0c9f822114f69c0c4c97ea1c120cf18f042a78bb","archive_status":"ready","archive_size":35009747703,"archive_retry_count":0,"records_status":"ready","archive_skip_reason":null,"zarr_errors":0,"zarr_failure_count":0,"zarr_deterministic":0,"zarr_failed_at":null,"num_dataset_citations":1,"num_datapaper_citations":6,"n_channels":2,"electrode_system":null,"has_hed":0,"hed_version":null,"is_exemplar":0,"bytes_present":40635239217,"data_complete":1,"withdrawn_at":null,"withdrawn_reason":null,"archive_complete":null,"archive_absent_files":null,"archive_declared_files":null,"zarr_pool_breaks":null,"total_recording_duration":204738.39599999998,"recording_duration_min":200.5,"recording_duration_max":8633.4,"recording_count":29,"recordings_unavailable":0,"recordings_measured":29,"channel_count_min":2,"channel_count_max":2,"sampling_frequency":10000,"power_line_frequency":60,"eeg_reference":"placed on FCz","placement_scheme":"based on the extended 10/20 system","sweep_stamps":"{\"enrichment_updated_at\":\"2026-08-19 01:42:27\",\"metadata_updated_at\":\"2026-08-19 01:42:42\",\"archive_checked_at\":\"2026-06-27 01:41:22\",\"zarr_checked_at\":null,\"records_checked_at\":\"2026-06-27 01:13:28\",\"citations_updated_at\":\"2026-09-08 03:00:50\",\"channel_montage_checked_at\":\"2026-06-28 23:42:47\",\"hed_checked_at\":\"2026-06-30 05:17:10\",\"data_checked_at\":\"2026-08-21 03:00:08\",\"availability_report_at\":\"2026-07-23 01:24:20\",\"recording_stats_at\":\"2026-09-02 11:33:18\",\"signal_defaults_at\":\"2026-09-02 12:34:47\"}","participants":25,"num_citations":7,"latest_version":"v1.0.0","zarr_verify_status":null,"zarr_verified_at":null,"owner_username":"nemarAdmin","owner_github":"nemarAdmin","file_size_formatted":"37.85 GB","zarr_data_failures":null,"zarr_index_url":"https://zarr.nemar.org/on005408/zarr/index.json","attestation_deposit_type":null,"attestation_key_status":null,"attestation_deidentified":null,"attestation_no_duplicate":null,"attestation_upstream_source":null,"attestation_accepted_at":null}}