{"dataset":{"id":"425","dataset_id":"nm000147","name":"The Brain, Body, and Behaviour Dataset (1.0.0) - Experiment 4","description":"The Brain, Body, and Behaviour Dataset (Experiment 4) is a multimodal neurophysiological dataset comprising simultaneous recordings of EEG, eye-tracking, cardiac, respiratory, and electrooculographic signals from 43 subjects across two sessions. Participants watched three educational videos (Stim-04, Stim-05, Stim-06) under two attention conditions. In Session 1 (attentive condition), participants viewed the videos and answered comprehension questions afterward. In Session 2 (distracted condition), participants viewed the same three videos in the same order while performing a concurrent backward counting task, with no comprehension testing. This derivative dataset supports investigation of neural and behavioral correlates of attention, learning, and cognitive load during naturalistic video viewing.","owner_user_id":19,"status":"active","github_repo":"nemarDatasets/nm000147","concept_doi":"10.82901/nemar.nm000147","latest_version_doi":"10.82901/nemar.nm000147.v1.0.0","created_at":"2026-04-13 21:28:55","updated_at":"2026-07-10 21:48:14","zenodo_concept_id":"20737917","is_sandbox":0,"visibility":"public","ezid_status":"public","enrichment_json":"{\n  \"version\": \"2.0\",\n  \"pipeline_stage\": \"enriched\",\n  \"title\": \"The Brain, Body, and Behaviour Dataset (1.0.0) - Experiment 4\",\n  \"description\": \"The Brain, Body, and Behaviour Dataset (Experiment 4) is a multimodal neurophysiological dataset comprising simultaneous recordings of EEG, eye-tracking, cardiac, respiratory, and electrooculographic signals from 43 subjects across two sessions. Participants watched three educational videos (Stim-04, Stim-05, Stim-06) under two attention conditions. In Session 1 (attentive condition), participants viewed the videos and answered comprehension questions afterward. In Session 2 (distracted condition), participants viewed the same three videos in the same order while performing a concurrent backward counting task, with no comprehension testing. This derivative dataset supports investigation of neural and behavioral correlates of attention, learning, and cognitive load during naturalistic video viewing.\",\n  \"methods_description\": \"Data collection involved simultaneous recording of EEG, EOG, ECG, respiration, pupillary responses, gaze coordinates, and head position during two sessions: Session 1 (attentive condition with post-video comprehension testing of three videos: Stim-04, Stim-05, Stim-06) and Session 2 (distracted condition with concurrent backward counting task while viewing the same three videos: Stim-04, Stim-05, Stim-06 in the same order, with no comprehension testing). EEG data were recorded in BDF format; physiological and eye-tracking data were converted to compressed TSV format using MATLAB.\",\n  \"license\": \"CC BY 4.0\",\n  \"dataset_type\": \"derivative\",\n  \"authors\": {\n    \"Jens Madsen\": {\n      \"orcid\": \"0000-0001-8163-001X\"\n    },\n    \"Nikhil Kuppa\": {},\n    \"Lucas Parra\": {}\n  },\n  \"keywords\": [\n    {\n      \"term\": \"EEG\"\n    },\n    {\n      \"term\": \"eye tracking\"\n    },\n    {\n      \"term\": \"attention\"\n    },\n    {\n      \"term\": \"learning\"\n    },\n    {\n      \"term\": \"multimodal neuroimaging\"\n    },\n    {\n      \"term\": \"cognitive load\"\n    },\n    {\n      \"term\": \"educational video\"\n    },\n    {\n      \"term\": \"physiological signals\"\n    },\n    {\n      \"term\": \"educational neuroscience\"\n    },\n    {\n      \"term\": \"distraction\"\n    },\n    {\n      \"term\": \"pupillometry\"\n    },\n    {\n      \"term\": \"ECG\"\n    },\n    {\n      \"term\": \"respiration\"\n    }\n  ],\n  \"related_identifiers\": [\n    {\n      \"identifier\": \"10.1101/2025.04.29.651259\",\n      \"identifier_type\": \"DOI\",\n      \"relation_type\": \"References\"\n    },\n    {\n      \"identifier\": \"https://github.com/nemarDatasets/nm000147\",\n      \"identifier_type\": \"URL\",\n      \"relation_type\": \"IsDescribedBy\"\n    },\n    {\n      \"identifier\": \"https://nemar.org/dataexplorer/detail?dataset_id=nm000147\",\n      \"identifier_type\": \"URL\",\n      \"relation_type\": \"IsDescribedBy\"\n    },\n    {\n      \"identifier\": \"10.1038/s41597-026-07215-1\",\n      \"identifier_type\": \"DOI\",\n      \"relation_type\": \"IsDerivedFrom\"\n    }\n  ],\n  \"funding_references\": [\n    {\n      \"funder_name\": \"National Science Foundation\",\n      \"award_number\": \"DRL-1660548\"\n    },\n    {\n      \"funder_name\": \"National Science Foundation\",\n      \"award_number\": \"DRL-2201835\"\n    }\n  ],\n  \"resource_type_general\": \"Dataset\",\n  \"resource_type_specific\": \"Neuroimaging Dataset\",\n  \"modalities\": [\n    \"beh\",\n    \"eeg\"\n  ],\n  \"sizes\": [\n    \"5.5 GB (5686 files)\"\n  ],\n  \"formats\": [\n    \".bdf\",\n    \".gz\",\n    \".json\",\n    \".md\",\n    \".sh\",\n    \".tsv\",\n    \".yml\"\n  ],\n  \"source_hash\": \"8d8eab4f8ca687b9991775664bd5916dccec705626348bdda7247caee0bb01d5\"\n}","last_activity_at":"2026-06-03 17:55:20","source":null,"source_id":null,"subject_count":null,"modalities":"beh,eeg","age_min":null,"age_max":null,"file_size":5499313593,"total_files":5686,"tasks":"stim01,stim02,stim03,stim04,stim05,stim06","metadata_columns_error":null,"staleness_warn_stage":null,"staleness_admin_notified_at":null,"authors":"Jens Madsen, Nikhil Kuppa, Lucas Parra","license":"CC BY 4.0","readme":"[![DOI](https://img.shields.io/badge/DOI-10.82901%2Fnemar.nm000147-blue)](https://doi.org/10.82901/nemar.nm000147)\n\n# The Brain, Body, and Behaviour Dataset - Experiment 4\n## Summary:\n\n**Description:**  Subjects watched five videos, knowing they'd be tested afterward. After each video, they answered 11 to 12 factual multiple-choice questions. Videos and questions were presented in random order.\n\n**Subjects:**  43, **Sessions:**  2 \n1. *Attentive* - Watch videos with focus (Stim 4, Stim 5, Stim 6)\n2. *Attentive* - Watch different videos (Stim 1, Stim 2, Stim 3) with focus and answer questions pertaining to them after\n\n**Tasks (Stimuli):**  \n\n-------------------------------------------------------------------------------------------------------------------------------------------\n| **Stimulus ID** | **Name**                                                    | **URL**                                                 |\n|-----------------|-------------------------------------------------------------|---------------------------------------------------------|\n| Stim-01         | Why are Stars Star-Shaped                                   | [Watch Here](https://www.youtube.com/embed/VVAKFJ8VVp4) |\n| Stim-02         | The Immune System Explained – Bacteria                      | [Watch Here](https://www.youtube.com/embed/zQGOcOUBi6s) |\n| Stim-03         | Are We All Related                                          | [Watch Here](https://www.youtube.com/embed/mnYSMhR3jCI) |\n| Stim-04         | How Modern Light Blbs Work                                  | [Watch Here](https://www.youtube.com/embed/oCEKMEeZXug) |\n| Stim-05         | What If We Killed All the Mosquitoes                        | [Watch Here](https://www.youtube.com/embed/9w-5wJYVmcw) |\n| Stim-06         | Three Factors That May Alter the Action of an Enzyme        | [Watch Here](https://www.youtube.com/embed/lkRZKqDdwzU) |\n-------------------------------------------------------------------------------------------------------------------------------------------\n\n**Modalities Recorded:**\n-   Gaze (X, Y)\n-   Pupil Size\n-   Blinks\n-   Saccades\n-   Head Position (X, Y, Z)\n-   EOG\n-   EEG\n-   ECG\n-   Respiration\n\n# Experiment Setup:\nIn experiment 4, there are 2 attend conditions where subjects were instructed to watch 3 instructional videos in attention condition, and 3 more videos in the second attention condition. \nThey were tested on the videos shown in the second attention condition. \n\nWe collected EEG, EOG, ECG, respiration, pupillary responses, gaze, and head position.\n\nSessions:\n\n-   Ses-01 contains these signals recorded on subjects watching these videos in an attentive condition. They answered questions pertaining to each video after watching the videos one at a time.\n-   Ses-02 contains the data recorded on subjects watching all of the videos again in a distracted condition in the same order as Ses-01 described above. The distraction from the stimuli was to silently count backwards from a random prime number in steps of 7. No questions were asked after this session.\n\n# Questionnaires:  \nQuestionnaires and answers to them can be found in the phenotype/ directory.\n\n1.  **stimuli_questionnaire:**\n\nThe stimuli_questionnaire tsv and json files have the questions, answers and correct answers.\n\n- “Domain” question type is a general domain knowledge question that was asked before the subject watched a video.\n- “Memory” question type is a memory testing question that is asked after the subject finishes watching the video, and has questions that are directly pertaining to the video content.\n\n2.  **asrs_questionnaire:**\n\nThese tsv and json files have questions and answers to the ASRS questionnaire that is an adult ADHD symptom checklist test. The answers are options from 0 (never) to 4 (very often).\n\n- There are two different scales used to score this test, and there are 2 different parts (Part A, Part B) to this test. Screen test involves just the first 6 questions (Part A), and the full-test involves the entire 18 question test (Part A + Part B). Despite lesser questions, the screen test (first 6 ques) gives one a higher indication of ADHD prevalence than the full-test.\n\n- This is the reason why the 2 different scoring scales are based on the Screen test. Scale one is out of 6, where each question counts for just 1 point depending on the frequency of the symptom occurrence, and if one scores over 4, there is a high chance of ADHD prevalence. The next scale is out of 24 where the 5 frequencies of symptom occurrence (never, rarely, sometimes, often, very often) are assigned scores in an increasing order of 0-4, and for each question, the respective scores of the frequency is the score for that question. Out of 24, the threshold for high prevalence of ADHD is 18 for this scale.\n\nThe following are the question numbers for inattentive ADHD and Hyperactive ADHD\n\ninattentive_questions = [1,2,3,4,7,8,9,10,11,12]\nhyperactive_questions = [5,6,13,14,15,16,17,18]\n\n# General BIDS dataset structure overview for all MEVD experiments\n\nEach BIDS dataset (one per experiment) has files that describe the dataset, its participants, and related metadata at the root directory - dataset_description.json, participants.tsv and participants.json, providing essential information about the study and participants to anyone working with the dataset.\n\nIn the root directory, the raw data of each participant is organized by subject (sub-XX) and then further divided into sessions (ses-XX) to accommodate multi-session data collection (some experiments have 2 sessions: attentive and distracted). Inside each session folder, you'll find modality-specific subfolders, such as:\n\n-   **eeg**: Contains electroencephalogram data files (.bdf), event logs (events.tsv), and additional metadata (.json) that describe the experiment and recording conditions.  \n    \n-   **beh**: Contains physiological recordings like ECG (electrocardiogram) and EOG (electrooculogram) stored in compressed .tsv.gz files, with accompanying metadata in .json files.  \n    \n-   **eyetrack**: Contains physiological recordings like eye-tracking (gaze coordinates and pupil size), and head movement data, stored in compressed .tsv.gz files, with accompanying metadata in .json files.\n-----------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------\n| Modality  | Filename Format                                                     | Data File Extension | Metadata                                     | Notes                                                                                      |\n|-----------|---------------------------------------------------------------------|---------------------|----------------------------------------------|--------------------------------------------------------------------------------------------|\n| EEG       | `sub_xx-ses_xx-task-stimxx_{file_of_interest.extension}`            | `.bdf`              | Exists for each file as a `.json` file       |  There are event files with a `.tsv` extension that include start and end times in seconds.|\n| Beh       | `sub_xx-ses_xx-task-stimxx_recording-{modality}_physio.{extension}` | `.tsv.gz`           | Exists for each file as a `.json` file       |                                                                                            |\n| EyeTrack  | `sub_xx-ses_xx-task-stimxx_{modality}_eyetrack.{extension}`         | `.tsv.gz`           | Exists for each file as a `.json` file       |                                                                                            |\n-----------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------------\n\nThere is also a **derivatives** directory which contains preprocessed data derived from the raw recordings, such as filtered heart rate data or preprocessed physiological signals, making it easy to ","bids_version":"1.10.0","sessions_count":2,"publish_date":"2026-04-13 21:28:55","embedding_dirty":0,"license_tier":"attribution","zarr_status":"ready","zarr_converted_at":"2026-09-06 02:13:37","zarr_store_count":256,"zarr_index_etag":"d587e71b0498417218793f06e3f7d427","zarr_source_commit":"56d4c56b2693775d81f88954f3d879d8346ae44d","archive_status":"ready","archive_size":4450170789,"archive_retry_count":0,"records_status":"ready","archive_skip_reason":null,"zarr_errors":0,"zarr_failure_count":0,"zarr_deterministic":0,"zarr_failed_at":null,"num_dataset_citations":0,"num_datapaper_citations":4,"n_channels":null,"electrode_system":null,"has_hed":0,"hed_version":null,"is_exemplar":0,"bytes_present":null,"data_complete":null,"withdrawn_at":null,"withdrawn_reason":null,"archive_complete":null,"archive_absent_files":null,"archive_declared_files":null,"zarr_pool_breaks":0,"total_recording_duration":68742,"recording_duration_min":143,"recording_duration_max":389,"recording_count":256,"recordings_unavailable":0,"recordings_measured":256,"channel_count_min":64,"channel_count_max":64,"sampling_frequency":null,"power_line_frequency":null,"eeg_reference":null,"placement_scheme":null,"sweep_stamps":"{\"enrichment_updated_at\":\"2026-06-17 20:22:12\",\"metadata_updated_at\":\"2026-06-17 20:22:57\",\"archive_checked_at\":\"2026-06-17 20:33:56\",\"zarr_checked_at\":null,\"records_checked_at\":\"2026-06-17 20:42:37\",\"citations_updated_at\":\"2026-09-08 03:00:50\",\"channel_montage_checked_at\":\"2026-06-28 22:52:39\",\"hed_checked_at\":\"2026-06-30 04:12:06\",\"data_checked_at\":null,\"availability_report_at\":\"2026-07-23 01:05:47\",\"signal_defaults_at\":\"2026-09-02 11:38:56\",\"recording_stats_at\":\"2026-09-06 03:00:35\",\"zarr_verify_attempted_at\":\"2026-09-07 03:02:47\"}","participants":0,"num_citations":4,"latest_version":"v1.0.0","zarr_verify_status":null,"zarr_verified_at":null,"owner_username":"bruaristimunha","owner_github":"bruAristimunha","file_size_formatted":"5.12 GB","zarr_data_failures":null,"zarr_index_url":"https://zarr.nemar.org/nm000147/zarr/index.json","attestation_deposit_type":null,"attestation_key_status":null,"attestation_deidentified":null,"attestation_no_duplicate":null,"attestation_upstream_source":null,"attestation_accepted_at":null}}