Child Language Longitudinal (ds003604)
A longitudinal neuroimaging dataset on language processing in children ages 5, 7, and 9 years old
Longitudinal brain MRI of 322 children from Austin, Texas, scanned at ages 5, 7 and 9 with T1w, 64-direction diffusion and four auditory language task fMRI scans, plus language and reading assessments, in BIDS under CC0.
Overview
This OpenNeuro dataset follows children through early school years to study how the brain specializes for spoken language. A lab at The University of Texas at Austin scanned children at about 5, 7 and 9 years of age while they judged word sounds, word meanings, sentence plausibility and sentence grammar. It was first released in April 2021, and the current snapshot 1.0.7 (November 2022) is published under a CC0 waiver.
Composition
The release holds 322 children who completed at least one structural and one functional scan: 174 girls and 148 boys. Two cohorts were recruited. The first (141 children) entered at 5.5 to 6.5 years (ses-5); 101 returned at 7 to 8 years (ses-7) and 46 of those again at 8.5 to 10 years (ses-9). The second (181 children) entered at ses-7, and 55 returned at ses-9. This gives 524 imaging sessions. Every child has T1-weighted and task fMRI data; 239 also have diffusion imaging. Standardized language, reading, IQ and articulation tests plus parent questionnaires are in the phenotype folder.
Acquisition
All scans were made on one Siemens Skyra 3T scanner with a 64-channel head coil. Structural images are 1 mm MPRAGE. Functional runs used multiband echo-planar imaging at 2 mm isotropic with a 1.25 s TR, two runs per task. Diffusion images have 64 directions at b = 800 s/mm². Phase-difference field maps were collected when time allowed. Children practised in a mock scanner first. T1-weighted images were defaced with pydeface using pediatric templates, and dates were shifted to protect identity.
Annotations
There are no image annotations. The release includes trial-level responses and reaction times for each fMRI run, per-run motion summaries, MRIQC reports, a composite T1 quality score and diffusion quality metrics from FSL eddy.
Known limitations
- Children with ADHD, psychiatric or neurological diagnoses, premature birth or substantial non-English language exposure were excluded, though recruitment also targeted children with language impairment.
- Not all children completed all tasks or sessions, and some sessions contain repeated scans.
- For 33 runs the original DICOMs were lost, so their JSON sidecars were reconstructed by hand.
- Image quality is somewhat lower than in datasets of older children because of the young age.
Cohort
Aggregate numbers from the sources below. Bars are relative to the 322 subjects.
Sex
- Female 174 54%
- Male 148 46%
Contrast combinations
How many subjects have exactly each set of contrasts.
| T1w | dwi | bold | Subjects with exactly this set |
|---|---|---|---|
239 | |||
83 |
Contrast / sequence
subjects, values can overlap
- T1-weighted 322 100%
- BOLD fMRI 322 100%
- Diffusion-weighted 239 74%
License and access
Our reading of the license, not legal advice. Before you use the data, read the original license and confirm that your use is allowed. We take no responsibility for how you use a dataset. Full disclaimer
Download without an account
Creative Commons Zero 1.0 Universal
Public domain dedication. Do anything with the data, including commercial use, without asking and without having to give credit.
dataset_description.json of snapshot 1.0.7 states "License" CC0; the paper's Data Records section also names CC0.
What you can do
- Yes
- Yes
- Yes
- Yes
What you can share
- Yes
- Yes
- Yes
What you must do
- No
- Share alike No
- No
- No
- Manuscript review No
- Release code No
- Return results No
- Delete after use No
Limits
- No
- Location limits No
Citation
Wang J, Lytle MN, Weiss Y, Yamasaki BL, Booth JR. A longitudinal neuroimaging dataset on language processing in children ages 5, 7, and 9 years old. Scientific Data 9, 4 (2022). doi:10.1038/s41597-021-01106-3
All numbers
Every number on this page, as stored in stats.csv, with its source.
| Measure | Breakdown | Value | Source |
|---|---|---|---|
| Subjects | total 322 subject folders in the snapshot 1.0.7 file tree; 322 rows in participants.tsv | 322 | wang2022 Methods: Participants |
| Subjects | sex=female sex Female | 174 | ds003604-participants sex |
| Subjects | sex=male sex Male | 148 | ds003604-participants sex |
| Subjects | contrast=T1w | 322 | ds003604-file-tree sub-*/ses-*/anat/*_T1w.nii.gz |
| Subjects | contrast=bold | 322 | ds003604-file-tree sub-*/ses-*/func/*_bold.nii.gz |
| Subjects | contrast=dwi | 239 | ds003604-file-tree sub-*/ses-*/dwi/*_dwi.nii.gz |
| Subjects | contrast_set=bold+dwi+T1w fieldmaps (phasediff) not counted as a contrast | 239 | ds003604-file-tree anat, func and dwi folders |
| Subjects | contrast_set=bold+T1w | 83 | ds003604-file-tree anat, func and dwi folders |
| Studies | total ses-5: 141, ses-7: 282, ses-9: 101; matches the cohort numbers in Background & Summary of the paper | 524 | ds003604-file-tree sub-*/ses-* folders |
| Scans | total all image files: 827 T1w, 3851 bold, 377 dwi and 365 phasediff field maps | 5,420 | ds003604-file-tree sub-*/ses-*/*/*.nii.gz |
| Scans | contrast=T1w includes repeated T1w scans within a session | 827 | ds003604-file-tree *_T1w.nii.gz |
| Scans | contrast=bold task fMRI runs | 3,851 | ds003604-file-tree *_bold.nii.gz |
| Scans | contrast=dwi | 377 | ds003604-file-tree *_dwi.nii.gz |
| Minimum age | total lower bound of the ses-5 enrollment age band (5.5 to 6.5 years) | 5.5 | wang2022 Background & Summary |
| Maximum age | total upper bound of the ses-9 age band (8.5 to 10 years) | 10 | wang2022 Background & Summary |
Sources
The keys used in the table above.
- wang2022 Wang et al. 2022, A longitudinal neuroimaging dataset on language processing in children ages 5, 7, and 9 years old, Scientific Data paper
- openneuro-ds003604-v107 OpenNeuro ds003604 snapshot 1.0.7 (dataset_description.json and snapshot summary) website
- ds003604-participants participants.tsv of ds003604 snapshot 1.0.7 (CC0), counted per row computed
- ds003604-file-tree File tree of ds003604 snapshot 1.0.7 (CC0), NIfTI files counted per subject, session and BIDS suffix computed