{"dataset":{"id":"60163","dataset_id":"on005170","name":"Chisco","description":"Chisco is a Chinese imagined speech EEG dataset collected from five participants, each recorded over 5-6 sessions. The dataset includes raw EEG recordings and preprocessed data (in fif and pkl formats), along with accompanying text stimuli used to elicit imagined speech. It is intended to support research on imagined speech decoding and brain-computer interface applications using EEG signals.","owner_user_id":15,"status":"active","github_repo":"nemarDatasets/on005170","concept_doi":"10.82901/nemar.on005170","latest_version_doi":"10.82901/nemar.on005170.v1.0.0","created_at":"2026-06-26 12:31:49","updated_at":"2026-08-19 01:59:12","zenodo_concept_id":null,"is_sandbox":0,"visibility":"public","ezid_status":"public","enrichment_json":"{\n  \"version\": \"2.0\",\n  \"pipeline_stage\": \"validated\",\n  \"title\": \"Chisco\",\n  \"description\": \"Chisco is a Chinese imagined speech EEG dataset collected from five participants, each recorded over 5-6 sessions. The dataset includes raw EEG recordings and preprocessed data (in fif and pkl formats), along with accompanying text stimuli used to elicit imagined speech. It is intended to support research on imagined speech decoding and brain-computer interface applications using EEG signals.\",\n  \"methods_description\": \"EEG data were collected from five participants (sub-01 to sub-05) across 5-6 sessions per subject, during which participants performed an imagined speech task guided by textual stimuli. Raw EEG data are provided in EDF format, with preprocessed derivatives available in fif and pkl formats.\",\n  \"license\": \"CC0\",\n  \"dataset_type\": \"raw\",\n  \"authors\": {\n    \"Zihan Zhang\": {},\n    \"Yi Zhao\": {},\n    \"Yu Bao\": {},\n    \"Xiao Ding\": {}\n  },\n  \"keywords\": [\n    {\n      \"term\": \"EEG\"\n    },\n    {\n      \"term\": \"imagined speech\"\n    },\n    {\n      \"term\": \"brain-computer interface\"\n    },\n    {\n      \"term\": \"Chinese language\"\n    },\n    {\n      \"term\": \"speech imagery\"\n    },\n    {\n      \"term\": \"BIDS\"\n    }\n  ],\n  \"related_identifiers\": [\n    {\n      \"identifier\": \"https://github.com/nemarDatasets/on005170\",\n      \"identifier_type\": \"URL\",\n      \"relation_type\": \"IsDescribedBy\"\n    },\n    {\n      \"identifier\": \"10.18112/openneuro.ds005170.v1.1.2\",\n      \"identifier_type\": \"DOI\",\n      \"relation_type\": \"IsDerivedFrom\"\n    },\n    {\n      \"identifier\": \"https://nemar.org/dataset/on005170\",\n      \"identifier_type\": \"URL\",\n      \"relation_type\": \"IsDescribedBy\"\n    }\n  ],\n  \"funding_references\": [\n    {\n      \"funder_name\": \"National Natural Science Foundation of China\",\n      \"award_number\": \"62176079\"\n    }\n  ],\n  \"resource_type_specific\": \"EEG Dataset\",\n  \"modalities\": [\n    \"eeg\"\n  ],\n  \"sizes\": [\n    \"280.6 GB (1167 files)\"\n  ],\n  \"formats\": [\n    \".edf\",\n    \".fif\",\n    \".json\",\n    \".md\",\n    \".pkl\",\n    \".tsv\",\n    \".xlsx\",\n    \".yml\"\n  ],\n  \"source_hash\": \"acb7e74ae5e9c1776961e99c841b6c541511722cc42c2db5e2db6efbbd2b8e1c\"\n}","last_activity_at":"2026-06-26 12:31:49","source":"openneuro","source_id":"ds005170","subject_count":5,"modalities":"eeg","age_min":22,"age_max":30,"file_size":281068719128,"total_files":1178,"tasks":"imagine,read","metadata_columns_error":null,"staleness_warn_stage":null,"staleness_admin_notified_at":null,"authors":"Zihan Zhang, Yi Zhao, Yu Bao, Xiao Ding","license":"CC0","readme":"[![DOI](https://img.shields.io/badge/DOI-10.82901%2Fnemar.on005170-blue)](https://doi.org/10.82901/nemar.on005170)\n\n# Chisco Dataset\n\nThis dataset is a Chinese imagined speech dataset with five participants, identified as sub-01 to sub-05. The dataset includes raw data and preprocessed data in both fif and pkl formats. Information also can be found in https://github.com/zhangzihan-is-good/Chisco\n\n## Supplementary Information\n\nThe initial dataset release encompassed data from three participants (sub-01 to sub-03) as detailed in related Chisco publications. Subsequently, data from two additional subjects (sub-04 and sub-05) were incorporated. During the interval between the original dataset release and the addition of the new data, the BIDS protocol underwent updates. To preserve the integrity of the data processing code presented in our publications, the supplementary data continue to adhere to the previous version of the BIDS protocol. Consequently, the BIDS validator on our website may report errors; however, these do not compromise the usability of the dataset.\n\nFuture releases will include data from sub-06 and sub-07, who participated under a new experimental paradigm. These will be published as part of a new dataset, Chisco 2.0. We invite you to stay tuned for further updates.\n\n## Dataset Structure\n\n### Root Directory\n\n- `dataset_description.json`\n- `participants.tsv`\n- `README`\n- `derivatives/`\n- `sub-01/` to `sub-05/`\n- `textdataset/`\n- `json/`\n\n### Raw Data\n\nThe root directory contains folders `sub-01` to `sub-05` with raw data. Each participant's folder contains 5-6 session folders, corresponding to data collected over 5-6 days.\n\n### Preprocessed Data\n\nPreprocessed data is stored in the `derivatives` folder in both fif and pkl formats.\n\n### Text Data\n\nThe `textdataset` folder and `json` folder contain text data used to stimulate the participants.\n\n### File Structure\n```\n/Chisco\n    /sub-01\n        /ses-01\n            /eeg\n                sub-01_ses-01_task-imagine_eeg.edf\n        ...\n    /sub-02\n        ...\n    /sub-03\n        ...\n    /derivatives\n        /fif\n            /sub-01\n                ...\n            /sub-02\n                ...\n            /sub-03\n                ...\n        /pkl\n            /sub-01\n                ...\n            /sub-02\n                ...\n            /sub-03\n                ...\n    /textdataset\n        ...\n    /json\n        ...\n    dataset_description.json\n    README\n    participants.tsv\n\n```\n## License\nThis dataset is licensed under the CC0 license. You are free to use the dataset for non-commercial purposes, but the original author needs to be properly indicated.\n\n## Citation\nIf you use this dataset in your research, please cite the following link:\n\nhttps://github.com/zhangzihan-is-good/Chisco\n\n## Contact Information\nFor any questions, please contact the dataset authors.\nThank you for using the Chisco!","bids_version":"1.6.0","sessions_count":6,"publish_date":null,"embedding_dirty":0,"license_tier":"public","zarr_status":"ready","zarr_converted_at":"2026-08-19 02:31:25","zarr_store_count":192,"zarr_index_etag":"dc8de73b647c76e2e01070e324ef4fff","zarr_source_commit":"12cddcd9d40c243eee5666c4bdde0b4b4f6a6411","archive_status":null,"archive_size":null,"archive_retry_count":0,"records_status":"ready","archive_skip_reason":"dataset 261.8 GB exceeds 100.0 GB archive limit; use direct download","zarr_errors":481,"zarr_failure_count":456,"zarr_deterministic":0,"zarr_failed_at":"2026-08-19 02:31:25","num_dataset_citations":0,"num_datapaper_citations":0,"n_channels":null,"electrode_system":null,"has_hed":0,"hed_version":null,"is_exemplar":0,"bytes_present":280568697258,"data_complete":1,"withdrawn_at":null,"withdrawn_reason":null,"archive_complete":null,"archive_absent_files":null,"archive_declared_files":null,"zarr_pool_breaks":null,"total_recording_duration":319229.4,"recording_duration_min":327,"recording_duration_max":2848.6,"recording_count":648,"recordings_unavailable":456,"recordings_measured":192,"channel_count_min":133,"channel_count_max":133,"sampling_frequency":null,"power_line_frequency":null,"eeg_reference":null,"placement_scheme":null,"sweep_stamps":"{\"enrichment_updated_at\":\"2026-08-19 01:59:01\",\"metadata_updated_at\":\"2026-08-19 01:59:11\",\"archive_checked_at\":\"2026-06-26 12:45:12\",\"zarr_checked_at\":null,\"records_checked_at\":\"2026-06-26 13:12:38\",\"citations_updated_at\":\"2026-09-08 03:00:54\",\"channel_montage_checked_at\":\"2026-06-28 23:38:47\",\"hed_checked_at\":\"2026-06-30 05:12:27\",\"data_checked_at\":\"2026-08-19 03:00:08\",\"availability_report_at\":\"2026-07-23 01:22:58\",\"recording_stats_at\":\"2026-09-02 11:33:11\",\"signal_defaults_at\":\"2026-09-02 12:25:33\"}","participants":5,"num_citations":0,"latest_version":"v1.0.0","zarr_verify_status":null,"zarr_verified_at":null,"owner_username":"nemarAdmin","owner_github":"nemarAdmin","file_size_formatted":"262 GB","zarr_data_failures":{"count":456,"detail_ref":"zarr/index.json","compacted_by":"migration_0074"},"zarr_index_url":"https://zarr.nemar.org/on005170/zarr/index.json","attestation_deposit_type":null,"attestation_key_status":null,"attestation_deidentified":null,"attestation_no_duplicate":null,"attestation_upstream_source":null,"attestation_accepted_at":null}}