{"dataset":{"id":"423","dataset_id":"nm000255","name":"The Brain, Body, and Behaviour Dataset (1.0.0) - Experiment 2","description":"The Brain, Body, and Behaviour Dataset (Experiment 2) is a multimodal neurophysiological dataset comprising 31 subjects across 2 sessions designed to investigate incidental learning and attentional modulation. Participants watched five educational videos under attentive and distracted conditions while simultaneous recordings of EEG, ECG, EOG, eye-tracking, and pupil dynamics were acquired. This derivative dataset includes preprocessed physiological signals and behavioral responses to memory questionnaires, providing a comprehensive resource for studying the neural and physiological correlates of attention, learning, and cognitive load.","owner_user_id":19,"status":"active","github_repo":"nemarDatasets/nm000255","concept_doi":"10.82901/nemar.nm000255","latest_version_doi":"10.82901/nemar.nm000255.v1.0.0","created_at":"2026-04-13 17:41:35","updated_at":"2026-07-10 22:01:59","zenodo_concept_id":"20522399","is_sandbox":0,"visibility":"public","ezid_status":"public","enrichment_json":"{\n  \"version\": \"2.0\",\n  \"pipeline_stage\": \"validated\",\n  \"authors\": {\n    \"Jens Madsen\": {},\n    \"Nikhil Kuppa\": {},\n    \"Lucas Parra\": {}\n  },\n  \"funding_references\": [\n    {\n      \"funder_name\": \"National Science Foundation\",\n      \"award_number\": \"DRL-1660548\"\n    },\n    {\n      \"funder_name\": \"National Science Foundation\",\n      \"award_number\": \"DRL-2201835\"\n    }\n  ],\n  \"title\": \"The Brain, Body, and Behaviour Dataset (1.0.0) - Experiment 2\",\n  \"license\": \"CC BY 4.0\",\n  \"dataset_type\": \"derivative\",\n  \"resource_type_general\": \"Dataset\",\n  \"modalities\": [\n    \"beh\",\n    \"eeg\"\n  ],\n  \"resource_type_specific\": \"Neuroimaging Dataset\",\n  \"related_identifiers\": [\n    {\n      \"identifier\": \"https://github.com/nemarDatasets/nm000255\",\n      \"identifier_type\": \"URL\",\n      \"relation_type\": \"IsDescribedBy\"\n    },\n    {\n      \"identifier\": \"https://nemar.org/dataexplorer/detail?dataset_id=nm000255\",\n      \"identifier_type\": \"URL\",\n      \"relation_type\": \"IsDescribedBy\"\n    },\n    {\n      \"identifier\": \"10.82901/nemar.nm000255\",\n      \"identifier_type\": \"DOI\",\n      \"relation_type\": \"IsVersionOf\"\n    }\n  ],\n  \"sizes\": [\n    \"5.6 GB (3800 files)\"\n  ],\n  \"formats\": [\n    \".bdf\",\n    \".gz\",\n    \".json\",\n    \".md\",\n    \".sh\",\n    \".tsv\",\n    \".yml\"\n  ],\n  \"description\": \"The Brain, Body, and Behaviour Dataset (Experiment 2) is a multimodal neurophysiological dataset comprising 31 subjects across 2 sessions designed to investigate incidental learning and attentional modulation. Participants watched five educational videos under attentive and distracted conditions while simultaneous recordings of EEG, ECG, EOG, eye-tracking, and pupil dynamics were acquired. This derivative dataset includes preprocessed physiological signals and behavioral responses to memory questionnaires, providing a comprehensive resource for studying the neural and physiological correlates of attention, learning, and cognitive load.\",\n  \"methods_description\": \"Data collection involved 31 subjects in two sessions: (1) Attentive condition—subjects watched five videos and answered 11-12 factual multiple-choice questions per video; (2) Distracted condition—subjects watched the same videos while silently counting backwards from a random prime number in steps of 7, with no subsequent testing. Simultaneous recordings included EEG (.bdf format), ECG, EOG, gaze coordinates (X, Y), pupil size, blinks, saccades, and head motion. Raw data were converted from .mat files to BIDS-compliant formats (.bdf for EEG, .tsv.gz for physiological and eye-tracking data) using MATLAB.\",\n  \"keywords\": [\n    {\n      \"term\": \"EEG\"\n    },\n    {\n      \"term\": \"eye tracking\"\n    },\n    {\n      \"term\": \"attention\"\n    },\n    {\n      \"term\": \"learning\"\n    },\n    {\n      \"term\": \"multimodal neuroimaging\"\n    },\n    {\n      \"term\": \"physiological signals\"\n    },\n    {\n      \"term\": \"incidental learning\"\n    }\n  ],\n  \"source_hash\": \"adf176082b6673aa31b92417b26fbb70c9b73491794acbadd1696a761c5635ca\"\n}","last_activity_at":"2026-04-13 18:55:19","source":null,"source_id":null,"subject_count":31,"modalities":"beh,eeg","age_min":18,"age_max":50,"file_size":5616916305,"total_files":3800,"tasks":"stim01,stim02,stim03,stim04,stim05","metadata_columns_error":null,"staleness_warn_stage":null,"staleness_admin_notified_at":null,"authors":"Jens Madsen, Nikhil Kuppa, Lucas Parra","license":"CC BY 4.0","readme":"[![DOI](https://img.shields.io/badge/DOI-10.82901%2Fnemar.nm000255-blue)](https://doi.org/10.82901/nemar.nm000255)\n\n# The Brain, Body, and Behaviour Dataset - Experiment 2\r\n## Summary:\r\n\r\n**Description:**  Subjects watched five videos, knowing they'd be tested afterward. After each video, they answered 11 to 12 factual multiple-choice questions. Videos and questions were presented in random order.\r\n\r\n**Subjects:**  31, **Sessions:**  2 \r\n1. *Attentive* - Watch videos with focus and answer questions after\r\n2. *Distracted* - Watch videos while counting backwards in your head, no test after watching\r\n\r\n# Tasks (Stimuli)\r\n\r\n## Experiment 2\r\n\r\n------------------------------------------------------------------------------------------------------------------------\r\n| **Stimulus ID** | **Name**                                 | **URL**                                                 |\r\n|-----------------|------------------------------------------|---------------------------------------------------------|\r\n| Stim-01         | Why are Stars Star-Shaped                | [Watch Here](https://www.youtube.com/embed/VVAKFJ8VVp4) |\r\n| Stim-02         | How Modern Light Bulbs Work              | [Watch Here](https://www.youtube.com/embed/oCEKMEeZXug) |\r\n| Stim-03         | The Immune System Explained – Bacteria   | [Watch Here](https://www.youtube.com/embed/zQGOcOUBi6s) |\r\n| Stim-04         | Who Invented the Internet - And Why      | [Watch Here](https://www.youtube.com/embed/21eFwbb48sE) |\r\n| Stim-05         | Why Do We Have More Boys Than Girls      | [Watch Here](https://www.youtube.com/embed/3IaYhG11ckA) |\r\n|----------------------------------------------------------------------------------------------------------------------|\r\n\r\n**Modalities Recorded:**\r\n-   Gaze (X, Y)\r\n-   Pupil Size\r\n-   Blinks\r\n-   Saccades\r\n-   EOG\r\n-   EEG\r\n-   ECG\r\n\r\n# Experiment Setup:\r\nIn experiment 2, subjects watched 5 different informative videos not knowing they would be questioned about the content of each video all together at the end of the video-watching. \r\nThis experiment was **incidental** learning because the subjects did not know they would be questioned about the contents of the video. We collected ECG, EEG, EOG, Head Motion, Gaze Coordinates, and Pupil Size.\r\n\r\nSessions:\r\n\r\n-   Ses-01 contains these signals recorded on subjects watching these videos in an attentive condition. They answered questions pertaining to each video after watching the videos one at a time.\r\n-   Ses-02 contains the data recorded on subjects watching all of the videos again in a distracted condition in the same order as Ses-01 described above. The distraction from the stimuli was to silently count backwards from a random prime number in steps of 7. No questions were asked after this session.\r\n\r\n# Questionnaires:  \r\nQuestionnaires and answers to them can be found in the phenotype/ directory.\r\n\r\n1.  **stimuli_questionnaire:**\r\n\r\nThe stimuli_questionnaire tsv and json files have the questions, answers and correct answers.\r\n\r\n- “Domain” question type is a general domain knowledge question that was asked before the subject watched a video.\r\n- “Memory” question type is a memory testing question that is asked after the subject finishes watching the video, and has questions that are directly pertaining to the video content.\r\n\r\n2.  **asrs_questionnaire:**\r\n\r\nThese tsv and json files have questions and answers to the ASRS questionnaire that is an adult ADHD symptom checklist test. The answers are options from 0 (never) to 4 (very often).\r\n\r\n- There are two different scales used to score this test, and there are 2 different parts (Part A, Part B) to this test. Screen test involves just the first 6 questions (Part A), and the full-test involves the entire 18 question test (Part A + Part B). Despite lesser questions, the screen test (first 6 ques) gives one a higher indication of ADHD prevalence than the full-test.\r\n\r\n- This is the reason why the 2 different scoring scales are based on the Screen test. Scale one is out of 6, where each question counts for just 1 point depending on the frequency of the symptom occurrence, and if one scores over 4, there is a high chance of ADHD prevalence. The next scale is out of 24 where the 5 frequencies of symptom occurrence (never, rarely, sometimes, often, very often) are assigned scores in an increasing order of 0-4, and for each question, the respective scores of the frequency is the score for that question. Out of 24, the threshold for high prevalence of ADHD is 18 for this scale.\r\n\r\nThe following are the question numbers for inattentive ADHD and Hyperactive ADHD\r\n\r\ninattentive_questions = [1,2,3,4,7,8,9,10,11,12]\r\nhyperactive_questions = [5,6,13,14,15,16,17,18]\r\n\r\n# General BIDS dataset structure overview for all MEVD experiments\r\n\r\nEach BIDS dataset (one per experiment) has files that describe the dataset, its participants, and related metadata at the root directory - dataset_description.json, participants.tsv and participants.json, providing essential information about the study and participants to anyone working with the dataset.\r\n\r\nIn the root directory, the raw data of each participant is organized by subject (sub-XX) and then further divided into sessions (ses-XX) to accommodate multi-session data collection (some experiments have 2 sessions: attentive and distracted). Inside each session folder, you'll find modality-specific subfolders, such as:\r\n\r\n-   **eeg**: Contains electroencephalogram data files (.bdf), event logs (events.tsv), and additional metadata (.json) that describe the experiment and recording conditions.  \r\n    \r\n-   **beh**: Contains physiological recordings like ECG (electrocardiogram) and EOG (electrooculogram) stored in compressed .tsv.gz files, with accompanying metadata in .json files.  \r\n    \r\n-   **eyetrack**: Contains physiological recordings like eye-tracking (gaze coordinates and pupil size), and head movement data, stored in compressed .tsv.gz files, with accompanying metadata in .json files.\r\n-----------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------\r\n| Modality  | Filename Format                                                     | Data File Extension | Metadata                                     | Notes                                                                                      |\r\n|-----------|---------------------------------------------------------------------|---------------------|----------------------------------------------|--------------------------------------------------------------------------------------------|\r\n| EEG       | `sub_xx-ses_xx-task-stimxx_{file_of_interest.extension}`            | `.bdf`              | Exists for each file as a `.json` file       |  There are event files with a `.tsv` extension that include start and end times in seconds.|\r\n| Beh       | `sub_xx-ses_xx-task-stimxx_recording-{modality}_physio.{extension}` | `.tsv.gz`           | Exists for each file as a `.json` file       |                                                                                            |\r\n| EyeTrack  | `sub_xx-ses_xx-task-stimxx_{modality}_eyetrack.{extension}`         | `.tsv.gz`           | Exists for each file as a `.json` file       |                                                                                            |\r\n-----------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------\r\n\r\nThere is also a **derivatives** directory which contains preprocessed data derived from the raw recordings, such as filtered heart rate data or preprocessed physiological signals, making it easy to work with and apply advanced analyses. Files in this directory are also stored in the BIDS structure (subject-wise → session-wise → modality-wise). A brief overview on what you can expect in the derivatives directory:\r\n\r\n-   **eeg**: Contains ","bids_version":"1.10.0","sessions_count":2,"publish_date":"2026-04-13 17:41:35","embedding_dirty":0,"license_tier":"attribution","zarr_status":"ready","zarr_converted_at":"2026-09-03 12:51:06","zarr_store_count":291,"zarr_index_etag":"fd61415d6c702d02cb389d7f54ab7113","zarr_source_commit":"290c8d6c5f2b0dca6a41f99127a479a5a9f73cd7","archive_status":"ready","archive_size":4591915246,"archive_retry_count":0,"records_status":null,"archive_skip_reason":null,"zarr_errors":0,"zarr_failure_count":0,"zarr_deterministic":0,"zarr_failed_at":null,"num_dataset_citations":0,"num_datapaper_citations":0,"n_channels":64,"electrode_system":"10-10","has_hed":0,"hed_version":null,"is_exemplar":0,"bytes_present":null,"data_complete":null,"withdrawn_at":null,"withdrawn_reason":null,"archive_complete":null,"archive_absent_files":null,"archive_declared_files":null,"zarr_pool_breaks":0,"total_recording_duration":71482,"recording_duration_min":143,"recording_duration_max":389,"recording_count":291,"recordings_unavailable":0,"recordings_measured":291,"channel_count_min":64,"channel_count_max":64,"sampling_frequency":128,"power_line_frequency":60,"eeg_reference":"none","placement_scheme":"10-20","sweep_stamps":"{\"enrichment_updated_at\":\"2026-06-03 07:13:54\",\"metadata_updated_at\":\"2026-06-03 07:16:56\",\"archive_checked_at\":\"2026-06-05 01:33:12\",\"zarr_checked_at\":\"2026-06-07 17:58:36\",\"records_checked_at\":null,\"citations_updated_at\":\"2026-08-17 03:01:00\",\"channel_montage_checked_at\":\"2026-06-28 23:01:01\",\"hed_checked_at\":\"2026-06-30 04:31:09\",\"data_checked_at\":null,\"availability_report_at\":\"2026-07-23 01:09:43\",\"signal_defaults_at\":\"2026-09-02 11:49:16\",\"recording_stats_at\":\"2026-09-04 03:00:23\"}","participants":31,"num_citations":0,"latest_version":"v1.0.0","zarr_verify_status":null,"zarr_verified_at":null,"owner_username":"bruaristimunha","owner_github":"bruAristimunha","file_size_formatted":"5.23 GB","zarr_data_failures":null,"zarr_index_url":"https://zarr.nemar.org/nm000255/zarr/index.json","attestation_deposit_type":null,"attestation_key_status":null,"attestation_deidentified":null,"attestation_no_duplicate":null,"attestation_upstream_source":null,"attestation_accepted_at":null}}