Aphasia Recovery Cohort (ARC)
Aphasia Recovery Cohort (ARC) Dataset
Longitudinal brain MRI of chronic left-hemisphere stroke survivors with aphasia, pooled from treatment and recovery studies to support lesion mapping and prediction of language impairment, with structural, diffusion, resting-state and naming-task scans, expert lesion masks and aphasia test scores.
Overview
The Aphasia Recovery Cohort brings together brain MRI, demographics and language test scores from people living with aphasia after a left-hemisphere stroke. The Aphasia Lab at the University of South Carolina pooled the data from several of its studies, among them anomia treatment, the POLAR outcome prediction protocol, speech entrainment and a brain stimulation trial. Because many people took part in more than one study, most of the cohort was scanned repeatedly. The data are shared on OpenNeuro in BIDS under CC0 and are meant for tool development, lesion mapping, impairment prediction and teaching.
Composition
The first release covers 230 people (89 female, 141 male) and 902 scanning sessions. All participants were in the chronic phase, at least six months after the stroke, and were 21 to 80 years old at enrolment. 141 people had a second session and 103 a third. Western Aphasia Battery scores and aphasia types are included; the paper lists 31 people without an aphasia diagnosis by the test cut-offs. participants.tsv holds 245 rows, more than the 230 people with images, and gives age at stroke rather than age at scan, so no age rows are listed here.
Acquisition
All scans come from Siemens 3 T scanners at one site, first Trio systems with 12-channel head coils and later Prisma systems with 20-channel head and neck coils. A first visit usually holds a 1 mm isotropic T1-weighted MP-RAGE and a T2-weighted SPACE scan. Other sessions add FLAIR, several diffusion protocols (some for tractography, some for kurtosis-type microstructure), resting-state fMRI and a sparse picture-naming fMRI task. Sequence parameters changed over the years and are stored in the BIDS sidecars. Anatomical images were defaced with spm_deface.
Annotations
An expert drew a lesion mask on the T2-weighted scan of each person's first visit. The snapshot holds 228 masks in derivatives/lesion_masks. The median lesion volume was 69.2 ml.
Known limitations
The cohort is retrospective and pools studies with different goals and protocols, so sequences vary between sessions. All participants volunteered for aphasia therapy, so lesions cluster in the middle cerebral artery territory and people with stroke but without aphasia are rare. No acute-phase imaging is included. The paper reports 441 T1-weighted and 447 T2-weighted series, while the file tree holds 447 and 441. The authors plan a second batch held back for competitions.
Cohort
Aggregate numbers from the sources below. Bars are relative to the 230 subjects.
Sex
- Female 89 39%
- Male 141 61%
Methods (Cohort) in Gibson et al. 2024, The Aphasia Recovery Cohort, an open-source chronic stroke repository, Scientific Data
Condition
Groups can overlap
- Stroke 230 100%
Methods (Cohort) in Gibson et al. 2024, The Aphasia Recovery Cohort, an open-source chronic stroke repository, Scientific Data
Field strength
- 3 T 230 100%
Methods (MRIs) in Gibson et al. 2024, The Aphasia Recovery Cohort, an open-source chronic stroke repository, Scientific Data
Contrast / sequence
Groups can overlap
- T1-weighted 229 100%
- T2-weighted 229 100%
- BOLD fMRI 220 96%
- Diffusion-weighted 217 94%
- FLAIR 138 60%
Data Records in Gibson et al. 2024, The Aphasia Recovery Cohort, an open-source chronic stroke repository, Scientific Data; From BIDS file tree of ds004884 snapshot 1.0.2 (CC0), raw NIfTI files under sub-*/ counted per suffix and lesion masks under derivatives/lesion_masks/
Scanner vendor
- Siemens Healthineers 230 100%
Methods (MRIs) in Gibson et al. 2024, The Aphasia Recovery Cohort, an open-source chronic stroke repository, Scientific Data
Country
- United States 230 100%
Methods (Cohort) in Gibson et al. 2024, The Aphasia Recovery Cohort, an open-source chronic stroke repository, Scientific Data
Contrast combinations
How many subjects have exactly each set of contrasts.
| T1w | T2w | FLAIR | dwi | bold | Subjects with exactly this set |
|---|---|---|---|---|---|
133 | |||||
79 | |||||
5 | |||||
4 | |||||
3 | |||||
3 | |||||
1 | |||||
1 | |||||
1 |
License and access
Our reading of the license, not legal advice. Before you use the data, read the original license and confirm that your use is allowed. We take no responsibility for how you use a dataset. Full disclaimer
Download without an account
Creative Commons Zero 1.0 Universal
Public domain dedication. Do anything with the data, including commercial use, without asking and without having to give credit.
dataset_description.json of snapshot 1.0.2 states "License" CC0.
What you can do
- Yes
- Yes
- Yes
- Yes
- Yes
What you can share
- Yes
- Yes
- Yes
What you must do
- No
- Share alike No
- No
- No
- Manuscript review No
- Release code No
- Return results No
- Delete after use No
Limits
- No
- Location limits No
Citation
Gibson M, Newman-Norlund R, Bonilha L, Fridriksson J, Hickok G, Hillis AE, den Ouden DB, Rorden C. The Aphasia Recovery Cohort, an open-source chronic stroke repository. Scientific Data 11, 981 (2024). doi:10.1038/s41597-024-03819-7
Sources
Every number on this page comes from one of these documents. Each chart names the table or page it is taken from. The raw numbers are in stats.csv.
- Gibson et al. 2024, The Aphasia Recovery Cohort, an open-source chronic stroke repository, Scientific Data paper
- OpenNeuro ds004884 snapshot 1.0.2 (dataset_description.json, CHANGES and snapshot size) website
- BIDS file tree of ds004884 snapshot 1.0.2 (CC0), raw NIfTI files under sub-*/ counted per suffix and lesion masks under derivatives/lesion_masks/ computed