{"dataset":{"id":"61150","dataset_id":"on006104","name":"EEG dataset for speech decoding","description":"This dataset comprises EEG recordings from 24 participants across two related studies (2019 and 2021) investigating phoneme discrimination during concurrent transcranial magnetic stimulation (TMS) of motor and speech-related cortical regions. Participants listened to speech sounds—including single phonemes, phoneme pairs, and phoneme triplets (real and pseudowords)—and responded via button press, enabling exploration of articulation and coarticulation effects on neural speech decoding. The dataset supports research into cortical mechanisms underlying speech perception and motor cortex involvement in phoneme processing.","owner_user_id":15,"status":"active","github_repo":"nemarDatasets/on006104","concept_doi":"10.82901/nemar.on006104","latest_version_doi":"10.82901/nemar.on006104.v1.0.0","created_at":"2026-06-28 13:00:59","updated_at":"2026-08-19 00:53:16","zenodo_concept_id":null,"is_sandbox":0,"visibility":"public","ezid_status":"public","enrichment_json":"{\n  \"version\": \"2.0\",\n  \"pipeline_stage\": \"validated\",\n  \"title\": \"EEG dataset for speech decoding\",\n  \"description\": \"This dataset comprises EEG recordings from 24 participants across two related studies (2019 and 2021) investigating phoneme discrimination during concurrent transcranial magnetic stimulation (TMS) of motor and speech-related cortical regions. Participants listened to speech sounds—including single phonemes, phoneme pairs, and phoneme triplets (real and pseudowords)—and responded via button press, enabling exploration of articulation and coarticulation effects on neural speech decoding. The dataset supports research into cortical mechanisms underlying speech perception and motor cortex involvement in phoneme processing.\",\n  \"methods_description\": \"EEG was recorded during a phoneme discrimination task in which participants listened to speech sounds and identified stimuli via button press. Stimuli included single consonants and vowels, CV/VC phoneme pairs, and CVC phoneme triplets (real and pseudowords). TMS was applied using a Magstim Super Rapid Plus1 stimulator with a figure-of-eight 40 mm coil, delivering paired pulses at 110% of resting motor threshold with a 50ms interpulse interval, targeting motor cortex regions (LipM1, TongueM1) in Study 1 and additional regions (Broca's area, verbal memory region) in Study 2.\",\n  \"license\": \"CC0\",\n  \"dataset_type\": \"raw\",\n  \"authors\": {\n    \"João Pedro Carvalho Moreira\": {},\n    \"Vinícius Rezende Carvalho\": {},\n    \"Eduardo Mazoni Andrade Marçal Mendes\": {},\n    \"Ariah Fallah\": {},\n    \"Terrence J. Sejnowski\": {},\n    \"Claudia Lainscsek\": {},\n    \"Lindy Comstock\": {}\n  },\n  \"keywords\": [\n    {\n      \"term\": \"EEG\"\n    },\n    {\n      \"term\": \"Transcranial Magnetic Stimulation\",\n      \"subject_scheme\": \"MeSH\",\n      \"value_uri\": \"http://id.nlm.nih.gov/mesh/D050781\"\n    },\n    {\n      \"term\": \"speech decoding\"\n    },\n    {\n      \"term\": \"phoneme discrimination\"\n    },\n    {\n      \"term\": \"Motor Cortex\",\n      \"subject_scheme\": \"MeSH\",\n      \"value_uri\": \"http://id.nlm.nih.gov/mesh/D009044\"\n    },\n    {\n      \"term\": \"coarticulation\"\n    },\n    {\n      \"term\": \"BIDS\"\n    }\n  ],\n  \"related_identifiers\": [\n    {\n      \"identifier\": \"https://github.com/nemarDatasets/on006104\",\n      \"identifier_type\": \"URL\",\n      \"relation_type\": \"IsDescribedBy\"\n    },\n    {\n      \"identifier\": \"https://nemar.org/dataset/on006104\",\n      \"identifier_type\": \"URL\",\n      \"relation_type\": \"IsDescribedBy\"\n    },\n    {\n      \"identifier\": \"10.1097/AUD.0b013e3181b1d42d\",\n      \"identifier_type\": \"DOI\",\n      \"relation_type\": \"References\"\n    },\n    {\n      \"identifier\": \"10.1523/JNEUROSCI.2383-16.2017\",\n      \"identifier_type\": \"DOI\",\n      \"relation_type\": \"References\"\n    },\n    {\n      \"identifier\": \"10.1002/hbm.22016\",\n      \"identifier_type\": \"DOI\",\n      \"relation_type\": \"References\"\n    },\n    {\n      \"identifier\": \"10.18112/openneuro.ds006104.v1.0.1\",\n      \"identifier_type\": \"DOI\",\n      \"relation_type\": \"IsDerivedFrom\"\n    }\n  ],\n  \"funding_references\": [\n    {\n      \"funder_name\": \"U.S. Russia Foundation\",\n      \"award_number\": \"20-AUG-19-UCLA\"\n    },\n    {\n      \"funder_name\": \"National Research University Higher School of Economics\"\n    }\n  ],\n  \"resource_type_general\": \"Dataset\",\n  \"resource_type_specific\": \"EEG Dataset\",\n  \"modalities\": [\n    \"eeg\"\n  ],\n  \"sizes\": [\n    \"90.0 GB (105 files)\"\n  ],\n  \"formats\": [\n    \".edf\",\n    \".fdt\",\n    \".json\",\n    \".md\",\n    \".set\",\n    \".tsv\",\n    \".yml\"\n  ],\n  \"source_hash\": \"ca73d825fc40117546119a92f211c1f652a015cc8909cc5f40fde24b70a880a4\"\n}","last_activity_at":"2026-06-28 13:00:59","source":"openneuro","source_id":"ds006104","subject_count":24,"modalities":"eeg","age_min":null,"age_max":null,"file_size":90006879537,"total_files":363,"tasks":"Words,phonemes,singlephoneme","metadata_columns_error":null,"staleness_warn_stage":null,"staleness_admin_notified_at":null,"authors":"João Pedro Carvalho Moreira, Vinícius Rezende Carvalho, Eduardo Mazoni Andrade Marçal Mendes, Ariah Fallah, Terrence J. Sejnowski, Claudia Lainscsek, Lindy Comstock","license":"CC0","readme":"[![DOI](https://img.shields.io/badge/DOI-10.82901%2Fnemar.on006104-blue)](https://doi.org/10.82901/nemar.on006104)\n\nEEG dataset for speech decoding\n============================\n\nDataset Overview\n---------------\n\nThis dataset contains EEG recordings from a phoneme discrimination task with TMS.\nThe data were collected during two related studies in 2019 and 2021.\n\nStudy 1 (2019, Session 01):\n- 8 participants (P01-P08)\n- Focus on CV and VC phoneme pairs\n- 2 blocks: CV pairs and VC pairs\n- TMS targeted to LipM1 (-56, -8, 46) and TongueM1 (-60, -10, 25)\n\nStudy 2 (2021, Session 02):\n- 16 participants (S01-S16)\n- Expanded to include single phonemes and phoneme triplets\n- 4 blocks: single phonemes, CV pairs, real words, and pseudowords\n- Additional TMS targets included Broca's area (BA 44: -51, 7, 23) and verbal memory region (BA 6: -46, 1, 41)\n\nTask Description\n---------------\n\nParticipants listened to speech sounds and identified stimuli with a button-press response.\nThe stimuli included:\n1. Single phonemes - Consonants (/b/, /p/, /d/, /t/, /s/, /z/) and vowels (/i/, /E/, /A/, /u/, /oU/)\n2. Phoneme pairs - CV and VC combinations of the phonemes\n3. Phoneme triplets - Real and pseudowords constructed of CVC sequences\n\nTMS Methodology\n--------------\n\nDetailed information about TMS parameters can be found in the sourcedata/tms_metadata/tms_parameters.json file.\nTMS was applied using a Magstim Super Rapid Plus1 stimulator with a figure-of-eight 40 mm coil.\nStimulation was delivered at 110% of resting motor threshold as paired pulses with 50ms interpulse interval.\n\nDetailed information about the methodology and results can be found in the associated publication:\nMoreira et al. \"An open-access EEG dataset for speech decoding: Exploring the role of articulation and coarticulation\"\n\n\n\nDirectory Structure\n------------------\n\nThe dataset follows BIDS convention with the following structure:\n/sub-[subject]/ses-[session]/eeg/\nWhere subject is P01-P08 for Study 1 and S01-S16 for Study 2.\nSession is 01 for Study 1 and 02 for Study 2.\n\nContact Information\n------------------\n\nFor questions about this dataset, please contact Lindy Comstock at lbcomstock@ucla.edu\n","bids_version":"1.6.0","sessions_count":2,"publish_date":null,"embedding_dirty":0,"license_tier":"public","zarr_status":"ready","zarr_converted_at":"2026-08-24 20:00:52","zarr_store_count":56,"zarr_index_etag":"6b0f4b2abc2b2c509168a9b4c4a2d678","zarr_source_commit":"2cb20338b2982b7c0126d8b9a5c24f6ccec9b9c5","archive_status":"ready","archive_size":51830138980,"archive_retry_count":0,"records_status":"ready","archive_skip_reason":null,"zarr_errors":0,"zarr_failure_count":0,"zarr_deterministic":0,"zarr_failed_at":null,"num_dataset_citations":6,"num_datapaper_citations":0,"n_channels":61,"electrode_system":"10-10","has_hed":0,"hed_version":null,"is_exemplar":0,"bytes_present":90004835624,"data_complete":1,"withdrawn_at":null,"withdrawn_reason":null,"archive_complete":null,"archive_absent_files":null,"archive_declared_files":null,"zarr_pool_breaks":0,"total_recording_duration":182725,"recording_duration_min":2389,"recording_duration_max":5865,"recording_count":56,"recordings_unavailable":0,"recordings_measured":56,"channel_count_min":62,"channel_count_max":84,"sampling_frequency":2000,"power_line_frequency":60,"eeg_reference":"CPz","placement_scheme":"extended 10-20 system","sweep_stamps":"{\"enrichment_updated_at\":\"2026-08-19 00:53:03\",\"metadata_updated_at\":\"2026-08-19 00:53:14\",\"archive_checked_at\":\"2026-06-28 13:51:27\",\"zarr_checked_at\":null,\"records_checked_at\":\"2026-06-28 13:21:30\",\"citations_updated_at\":\"2026-09-05 03:00:22\",\"channel_montage_checked_at\":\"2026-06-28 23:53:29\",\"hed_checked_at\":\"2026-06-30 05:30:42\",\"data_checked_at\":\"2026-08-26 03:00:25\",\"availability_report_at\":\"2026-07-23 01:28:21\",\"recording_stats_at\":\"2026-09-02 11:33:35\",\"signal_defaults_at\":\"2026-09-02 12:43:17\"}","participants":24,"num_citations":6,"latest_version":"v1.0.0","zarr_verify_status":null,"zarr_verified_at":null,"owner_username":"nemarAdmin","owner_github":"nemarAdmin","file_size_formatted":"83.83 GB","zarr_data_failures":null,"zarr_index_url":"https://zarr.nemar.org/on006104/zarr/index.json","attestation_deposit_type":null,"attestation_key_status":null,"attestation_deidentified":null,"attestation_no_duplicate":null,"attestation_upstream_source":null,"attestation_accepted_at":null}}