Abstract
Prior work in late-onset Alzheimer’s disease (LOAD) has resulted in discrepant findings as to whether recent consanguinity and outbred autozygosity are associated with LOAD risk. In the current study, we tested the association between consanguinity and outbred autozygosity with LOAD in the largest such analysis to date, in which 20 LOAD GWAS datasets were retrieved through public databases. Our analyses were restricted to eight distinct ethnic groups: African–Caribbean, Ashkenazi–Jewish European, European–Caribbean, French–Canadian, Finnish European, North-Western European, South-Eastern European, and Yoruba African for a total of 21,492 unrelated subjects (11,196 LOAD and 10,296 controls). Recent consanguinity determination was performed using FSuite v1.0.3, according to subjects’ ancestral background. The level of autozygosity in the outbred population was assessed by calculating inbreeding estimates based on the proportion (FROH) and the number (NROH) of runs of homozygosity (ROHs). We analyzed all eight ethnic groups using a fixed-effect meta-analysis, which showed a significant association of recent consanguinity with LOAD (N = 21,481; OR = 1.262, P = 3.6 × 10–4), independently of APOE∗4 (N = 21,468, OR = 1.237, P = 0.002), and years of education (N = 9,257; OR = 1.274, P = 0.020). Autozygosity in the outbred population was also associated with an increased risk of LOAD, both for FROH (N = 20,237; OR = 1.204, P = 0.030) and NROH metrics (N = 20,237; OR = 1.019, P = 0.006), independently of APOE∗4 [(FROH, N = 20,225; OR = 1.222, P = 0.029) (NROH, N = 20,225; OR = 1.019, P = 0.007)]. By leveraging the Alzheimer’s Disease Sequencing Project (ADSP) whole-exome sequencing (WES) data, we determined that LOAD subjects do not show an enrichment of rare, risk-enhancing minor homozygote variants compared to the control population. A two-stage recessive GWAS using ADSP data from 201 consanguineous subjects in the discovery phase followed by validation in 10,469 subjects led to the identification of RPH3AL p.A303V (rs117190076) as a rare minor homozygote variant increasing the risk of LOAD [discovery: Genotype Relative Risk (GRR) = 46, P = 2.16 × 10–6; validation: GRR = 1.9, P = 8.0 × 10–4]. These results confirm that recent consanguinity and autozygosity in the outbred population increase risk for LOAD. Subsequent work, with increased samples sizes of consanguineous subjects, should accelerate the discovery of non-additive genetic effects in LOAD.
Introduction
The impact of consanguinity on reproduction and Mendelian disorders is well known and documented (). In contrast, very little has been published on the effects of consanguinity on late-onset diseases, even though inbreeding may have a prominent influence on late-onset traits (). Recessive inheritance of complex phenotypes can be linked to long [≥1-megabase (Mb)] runs of homozygosity (ROHs), which are indicative of recent consanguinity (). Levels of homozygosity vary by population owing to the evolutionary distance of different populations from the ancient migration events that led to elevated homozygosity (; ).
Several studies have been carried out in late-onset Alzheimer’s Disease (LOAD) cohorts from different ethnicities, including Caribbean-Hispanics (), African Americans (), Wadi-Ara Arabs (), and Northern-Europeans (; ) with the aim of determining the impact of ROHs on LOAD. The studies carried out in Caribbean-Hispanic () and African American () populations both demonstrated an association of long ROHs with LOAD, thus suggesting a link between recent consanguinity and LOAD. An association between consanguinity and LOAD was also demonstrated in a genealogical study of the Saguenay region in Québec (). Conversely, in the small ethnic isolate of Wadi-Ara Arabs (), the average degree of inbreeding was significantly higher in controls compared to cases. Moreover, the two studies carried out in Caucasians (; ) showed discordant findings: the British–Irish study () displayed no association of number of ROHs with LOAD, while the mainly North-Western European cohort of neuropathologically verified subjects from the TGenII cohort () showed a suggestive increased number of ROHs in LOAD cases compared to controls. In sum, the results vary considerably by ancestral background, thus failing to provide a clear picture of the overall impact of homozygosity on LOAD risk.
In the present work we tested the association of consanguinity and autozygosity with LOAD by leveraging a large collection of publicly available GWAS data. To this aim, we determined the individual ancestry of subjects belonging to 20 independent GWAS datasets and pooled the consanguineous subjects according to their respective ethnic group. This step was followed by an association analysis between consanguinity in LOAD cases against older, cognitively healthy controls. We also tested the overall impact of genome-wide autozygosity in the outbred population. Finally, we leveraged Whole-Exome Sequencing (WES) data from the Alzheimer’s Disease Sequencing Project (ADSP) () to test the global burden of rare minor homozygote variants in LOAD and to perform a two-stage recessive GWAS using 201 consanguineous subjects in the discovery phase followed by validation in 10,469 subjects.
Materials and Methods
Subjects
Twenty LOAD GWAS datasets were obtained from publicly available data repositories (Supplementary Table S1). The 20 datasets have been described in previous studies (; ; ; ; ; ; ; ). Details of the participating studies and genotyping platforms used are provided in Supplementary Table S1.
The Alzheimer’s Disease Neuroimaging Initiative (ADNI) () was launched in 2003 as a public-private partnership, led by Principal Investigator Michael W. Weiner, MD. The primary goal of ADNI has been to test whether serial magnetic resonance imaging, positron emission tomography, other biological markers, and clinical and neuropsychological assessment can be combined to measure the progression of mild cognitive impairment and early Alzheimer’s disease.
Whole-exome sequencing from the discovery phase of the Alzheimer’s Disease Sequencing Project (ADSP) () was obtained through the National Institute on Aging Genetics of Alzheimer’s Disease Data Storage Site (NIAGADS) and it includes 5,096 Alzheimer’s Disease (AD) cases and 4,965 controls, with an additional enriched sample set comprised of 853 AD cases from multiple affected families and 171 Hispanic controls.
This was a re-analysis of de-identified data available from shared data repositories. The study protocol was granted an exemption by the Stanford Institutional Review Board because the analyses were carried out on “de-identified, off-the-shelf” data.
Inclusion Criteria, Quality Control (QC) Pipeline, Ancestry Determination and Imputation
The entire dataset includes 34,111 participants. Analyses were performed using PLINK 1.9 (). A comprehensive flowchart of the data QC/harmonization/ancestry-determination steps applied to the full dataset is reported as Figure 1.
FIGURE 1
Subjects with autosome missingness (≥5%) and/or X-chromosome missingness (≥5%) within the same dataset, age below 60 years, age information missing, or phenotype inconsistency [missing phenotype, diagnosis of mild cognitive impairment or other neurodegenerative phenotype] were excluded from the analysis (Supplementary Table S2).
Individual ancestry was determined using SNPweights v.2.1 () using reference populations from the 1000 Genomes Project (1KGP) (). By applying an ancestry percentage cut-off ≥ 80%, the samples were stratified into the five super populations, South-Asians (SAS), East-Asians (EAS), Americans (AMR), Africans (AFR), and Europeans (EUR) (Supplementary Table S2). Since most of the samples belonged to the European population, we also determined their ancestry percentage according to four major ethnicities, North-Western, South-Eastern, Ashkenazi-Jewish, and Finnish Europeans, using reference populations available both from SNPweights v.2.1 () and 1KGP (). European subjects were stratified into the above-mentioned ethnicities when their ancestry percentage was attributable with an ancestry percentage cut-off ≥ 50% (Supplementary Table S3).
We assigned French–Canadian (FCN) ancestry to subjects included in GenADA GWAS () when they both reported Canada as country of origin for all the grandparents and French as their first spoken language.
Most subjects belonging to the Columbia University Study of Caribbean Hispanics with Familial and Sporadic Late Onset Alzheimer’s disease (CIDR) () had admixed ancestry (Supplementary Table S2); therefore, we stratified the subjects into three groups according to their prevalent ancestral background. This stratification allowed the definition of one dataset composed of African–Caribbean (African ancestry ≥ 50%), one dataset composed of European–Caribbean (European ancestry ≥ 50%) and one dataset including highly admixed subjects (ancestry percentage less than 50% attributable to a unique super-population from 1KGP, ). Only the African–Caribbean and the European–Caribbean datasets were considered for the analyses. Subjects with genetic ancestry estimates discordant from self-reported ancestry were excluded from the analyses.
Next, datasets were tested for presence of consanguineous subjects using FSuite v.1.0.3 (). Consanguineous female subjects were flagged to avoid their exclusion because of apparent sex-inconsistency (e.g., due to increased homozygosity at X-chromosome SNPs). Subjects showing sex-inconsistency were excluded along with the possibly contaminated samples {heterozygosity F ≤ −0.03; more than 25 related [(identity-by-descent (IBD) ≥ 0.0625, equivalent to 3rd degree relative] within the same dataset}.
With the aim of maximizing the efficiency of quality control procedures and harmonizing the GWAS results after imputation, we collapsed the genotyping data from the 20 GWAS (when needed), according to sample ethnicity and the number of SNPs shared across the SNP-array platforms. Thus, we defined five groups, reported in Supplementary Table S4, where the subjects from different GWAS were collapsed and further QCed to remove the SNPs with a call rate ≤ 95%; Minor Allele Frequency (MAF) ≤ 1%; SNPs with MAF deviating more than 10% from the MAF reported in 1KGP for the relative population; SNPs with differential missingness between cases and controls (P < 0.05); SNPs deviating from Hardy–Weinberg Equilibrium (HWE) in controls (P < 5 × 10–5); tri-allelic SNPs; and SNPs where the alleles are mismatched compared to the 1KGP reference sequence. A/T and C/G SNPs were removed prior to imputation.
All the datasets were phased and imputed using the Michigan Imputation Server (), considering the Haplotype Reference Consortium r1.1. 2016 European panel () for Europeans, 1KGP Phase 3 African panel () for African Indianapolis-Ibadan and the Consortium on Asthma among African-ancestry Populations in the Americas (CAAPA) () for admixed Caribbean. After imputation, SNPs with a r2 quality score ≤ 0.7 and MAF ≤ 0.01 were excluded. Results of the imputation process are provided in Supplementary Table S5 and Supplementary Figure S1.
For the statistical analyses, inter-dataset duplicates (IBD ≥ 0.95) were removed from the dataset having the lowest SNP coverage, while, in case of relatedness (IBD ≥ 0.0625) the affected or older subject were kept, independently of SNP coverage (Supplementary Tables S6–S13).
Consanguinity Determination
Consanguinity determination was performed using FSuite v1.0.3 (), both at the pre- and post-imputation stage on QCed SNPs, according to each subjects’ ancestral background. Results were concordant in 98% of the subjects analyzed. Discordant subjects were kept in the subsequent analyses as outbred, since their ROHs may be somatic Copy Number Variations or linked to a specific ancestral background (e.g., ethnic minorities – Acadians, Sardinians, etc.) not captured by our ancestry-determination pipeline.
Subjects showing a homozygous region over 10 Mb on only one chromosome were considered carriers of putative uniparental isodisomy (UPD), according to the homozygosity cut-off previously reported (). UPD carriers were excluded from the association testing of consanguinity with LOAD since they represent subjects affected by chromosomal alterations, not the result of consanguineous unions.
ROH Calling and Burden Analysis
Runs of homozygosity (≥1 Mb) were determined for each ethnic group separately using PLINK1.9 () and according to the guidelines recently reported (). GWAS datasets were pruned for strong LD (MAF ≥ 0.01, r2 ≤ 0.1) and ROHs were defined as being ≥ 65 consecutive homozygous SNPs with no heterozygote calls allowed and a density greater than 1 SNP per 200 kb.
For each subject, we summed the total length of all their ROHs in the autosome and divided by the total SNP-mappable autosomal distance (2.77 × 109 bases) to derive FROH, the proportion (between 0 and 1) of the autosome in ROHs, as previously described (). FROH was used as the predictor of case-control status in ROH burden analyses.
Population Genetics, WES Variant Annotation and Statistical Analyses
Inter-population structure was examined using ADMIXTURE (). Intra-population structure for each ancestral group was determined using principal components (PCs) obtained from EIGENSOFT v.6.1 () using pruned (r2 ≤ 0.1), directly genotyped, SNPs.
The association of consanguinity/autozygosity with LOAD was carried out using three different logistic regression models:
- (a)
MODEL1: adjusting for each subject’s age at LOAD onset (when not available, we used the subject’s age at last visit or death), sex, the first three PC eigenvalues from population structure and GWAS imputation group (only for Ashkenazi-Jewish Europeans, Finnish Europeans, North-Western Europeans and South-Eastern Europeans, since those subjects were genotyped on multiple platforms);
- (b)
MODEL2: including APOE∗4 dose to the list of covariates used in MODEL1;
- (c)
MODEL3: including “years of education” (EDU) to the list of covariates used in MODEL2.
Since EDU is available only for 43.1% of the full dataset (9,260 out of 21,492 subjects), we fitted MODEL1 and MODEL2 to this restricted subset of subjects as well, to allow an appropriate comparison with MODEL3 of the effect estimates. The association of EDU with FROH was tested by linear regression, considering MODEL2 and adjusting for diagnosis status. The analyses were carried out for each ethnic group separately and combined using a fixed-effect meta-analysis implemented in GWAMA ().
GWAS participants were mapped to ADSP WES participants through IBD estimation using ∼30K common, overlapping SNPs (MAF > 0.01) between the two datasets. We considered matching samples as those with a pair-wise coefficient of relatedness above 99%. The burden of rare recessive variants was tested using a general linear model, adjusting for sex, age, APOE∗4 dose and ethnicity. ADSP WES data were annotated using Variant Effect Predictor (VEP) (). Recessive variants having a CADD () score ≥ 15 and a MAF < 10% were considered when testing the global burden of rare recessive variants in LOAD and in a two-stage GWAS by applying the Recessive-Allele Frequency Test (RAFT) statistic ().
Statistical significance was set at P < 0.05 for all the association testing, while for the two-stage GWAS, we applied Bonferroni’s correction according to the number of recessive variants tested in the discovery [P < 1.8 × 10–5 (0.05/2767)] and replication [P < 0.007 (0.05/7)] phases.
Results
Ancestry Determination and Ethnic-Specific Differences
After quality control procedures, our analyses were restricted to eight distinct ethnic groups, namely African–Caribbeans from Dominican Republic (ACD, N = 398), Ashkenazi-Jewish Europeans (AJE, N = 1,229), European–Caribbeans from Dominican Republic (ECD, N = 671), French–Canadians (FCN, N = 376), Finnish Europeans (FIN, N = 219), North-Western Europeans (NWE, N = 16,496), South-Eastern Europeans (SEE, N = 1,083), and African-Yoruba (YRI, N = 1,020), for a total of 21,492 unrelated subjects (11,196 LOAD, 10,296 controls). The analysis of genetic structure, applied to the whole dataset, confirmed the effectiveness of ancestry determination cut-offs (Figure 2).
FIGURE 2
Ethnic groups showed significant differences (P < 0.00001) in mean age, EDU and APOE allele frequency, independently of diagnostic status (Figure 3). Both consanguinity rates and consanguinity degree differed significantly (P < 0.00001) across the eight ethnic groups analyzed, with percentages of consanguineous subjects ranging from 1.2% in YRI to 31.7% in ECD (Table 1). As expected, consanguineous subjects displayed higher FROH estimates (0.024 ± 0.025 vs. 0.003 ± 0.002) and more ROH (NROH, 9.4 ± 6.3 vs. 4.4 ± 2.5) compared to outbred subjects (P < 0.00001, Table 2). When considering exclusively the outbred population, both FROH estimates and NROH significantly differed across the eight ethnic groups analyzed (P < 0.00001, Figure 3), with ACD and ECD showing the lowest FROH (0.0007 ± 0.0012) and NROH (0.4 ± 0.8), respectively. Conversely, the highest FROH and NROH were found in FIN (0.0053 ± 0.0038) and FCN (8.4 ± 2.9), respectively (Figure 3).
FIGURE 3

The eight ethnic groups analyzed showed ethnic-specific differences in Alzheimer’s-relevant risk factors. (A) mean age, (B) EDU, (C)APOE∗2 frequency, (D)APOE∗4 frequency, (E) inbreeding coefficient, and (F) number of ROHs (>1 Mb) differ significantly across the eight ethnic groups analyzed (P < 0.00001).
TABLE 1
| Ethnicity | Status | N | INBRED N (%) | 1C N (INBRED-%) | 2C N (INBRED-%) | 2 × 1C N (INBRED-%) | AV N (INBRED-%) |
| ACD | Total | 398 | 94 (23.6) | 13 (13.8) | 80 (85.1) | — | 1 (1.1) |
| LOAD | 170 | 47 (27.7) | 10 (21.3) | 37 (81.2) | — | — | |
| Controls | 228 | 47 (20.6) | 3 (6.4) | 43 (91.5) | — | 1 (2.1) | |
| AJE | Total | 1229 | 141 (11.5) | 34 (24.1) | 107 (75.8) | — | 1 (0.1) |
| LOAD | 754 | 80 (10.6) | 21 (26.3) | 59 (73.8) | — | — | |
| Controls | 475 | 61 (12.8) | 13 (21.3) | 48 (78.7) | — | 1 (1.6) | |
| ECD | Total | 671 | 213 (31.7) | 49 (23.0) | 158 (74.2) | 5 (2.3) | 1 (0.5) |
| LOAD | 337 | 117 (34.7) | 32 (27.4) | 80 (68.4) | 4 (3.4) | 1 (0.8) | |
| Controls | 334 | 96 (28.7) | 17 (17.7) | 78 (81.3) | 1 (1.0) | — | |
| FCN | Total | 376 | 53 (14.1) | 3 (5.7) | 50 (94.3) | — | — |
| LOAD | 193 | 29 (15.0) | 1 (3.4) | 28 (96.6) | — | — | |
| Controls | 183 | 24 (13.1) | 2 (8.3) | 22 (91.7) | — | — | |
| FIN | Total | 219 | 31 (14.2) | — | 31 (100.0) | — | |
| LOAD | 110 | 18 (16.4) | — | 18 (100.0) | — | — | |
| Controls | 109 | 13 (11.9) | — | 13 (100.0) | — | — | |
| NWE | Total | 16,496 | 500 (3.0) | 55 (11.0) | 439 (87.8) | 6 (1.2) | — |
| LOAD | 8,870 | 313 (3.5) | 37 (11.8) | 270 (86.3) | 6 (1.9) | — | |
| Controls | 7,626 | 187 (2.5) | 18 (9.6) | 169 (90.4) | — | — | |
| SEE | Total | 1,083 | 184 (17.0) | 13 (7.1) | 171 (92.9) | — | — |
| LOAD | 678 | 116 (17.1) | 10 (8.6) | 106 (91.4) | — | — | |
| Controls | 405 | 68 (16.8) | 3 (4.4) | 65 (95.6) | — | — | |
| YRI | Total | 1,020 | 12 (1.2) | — | 12 (100.0) | — | — |
| LOAD | 84 | 1 (1.2) | — | 1 (100.0) | — | — | |
| Controls | 936 | 11 (1.2) | — | 11 (100.0) | — | — |
Consanguinity rates differ across the eight ethnic groups analyzed.
1C, first-cousins offspring; 2C, second-cousins offspring; 2 × 1C, double-first cousins offspring; AV, avuncular offspring. ACD, African–Caribbean from Dominican Republic; AJE, Ashkenazi-Jewish Europeans; ECD, European–Caribbean from Dominican Republic; FCN, French–Canadians; FIN, Finnish Europeans; NWE, North-Western Europeans; SEE, South-Eastern Europeans; YRI, African Yoruba.
TABLE 2
| Ethnicity | Status | Outbred FROH (Mean ± SD) | Inbred FROH (Mean ± SD) | Outbred ROH N (Mean ± SD) | Inbred ROH N (Mean ± SD) |
| ACD | Total | 0.0007 ± 0.0012 | 0.0251 ± 0.0328 | 0.6 ± 0.8 | 6.6 ± 5.5 |
| LOAD | 0.0006 ± 0.0011 | 0.0280 ± 0.0233 | 0.6 ± 0.8 | 7.8 ± 5.6 | |
| Controls | 0.0007 ± 0.0012 | 0.0223 ± 0.0402 | 0.6 ± 0.8 | 5.4 ± 5.1 | |
| AJE | Total | 0.0048 ± 0.0031 | 0.0286 ± 0.0255 | 3.4 ± 2.0 | 9.3 ± 5.4 |
| LOAD | 0.0048 ± 0.0032 | 0.0287 ± 0.0233 | 3.4 ± 2.0 | 9.0 ± 5.3 | |
| Controls | 0.0049 ± 0.0030 | 0.0285 ± 0.0284 | 3.6 ± 2.0 | 9.6 ± 5.7 | |
| ECD | Total | 0.0008 ± 0.0015 | 0.0275 ± 0.0303 | 0.4 ± 0.8 | 6.9 ± 5.8 |
| LOAD | 0.0007 ± 0.0016 | 0.0313 ± 0.0349 | 0.4 ± 0.8 | 7.9 ± 6.6 | |
| Controls | 0.0007 ± 0.0015 | 0.0229 ± 0.0228 | 0.4 ± 0.8 | 5.8 ± 4.5 | |
| FCN | Total | 0.0052 ± 0.0022 | 0.0227 ± 0.0154 | 8.4 ± 2.9 | 13.9 ± 5.1 |
| LOAD | 0.0054 ± 0.0022 | 0.0238 ± 0.0165 | 8.8 ± 3.0 | 14.6 ± 5.3 | |
| Controls | 0.0050 ± 0.0023 | 0.0215 ± 0.0141 | 8.0 ± 2.9 | 13.0 ± 4.9 | |
| FIN | Total | 0.0053 ± 0.0038 | 0.0156 ± 0.0067 | 5.8 ± 3.3 | 10.8 ± 3.3 |
| LOAD | 0.0052 ± 0.0036 | 0.0162 ± 0.0077 | 5.7 ± 3.2 | 11.1 ± 3.5 | |
| Controls | 0.0053 ± 0.0040 | 0.0148 ± 0.0054 | 5.8 ± 3.5 | 10.4 ± 3.1 | |
| NWE | Total | 0.0029 ± 0.0016 | 0.0233 ± 0.0233 | 4.7 ± 2.3 | 11.3 ± 6.7 |
| LOAD | 0.0029 ± 0.0016 | 0.0242 ± 0.0264 | 4.8 ± 2.3 | 11.4 ± 7.2 | |
| Controls | 0.0029 ± 0.0016 | 0.0219 ± 0.0173 | 4.7 ± 2.3 | 11.2 ± 5.9 | |
| SEE | Total | 0.0013 ± 0.0020 | 0.0198 ± 0.0182 | 1.3 ± 1.6 | 7.9 ± 5.2 |
| LOAD | 0.0014 ± 0.0020 | 0.0200 ± 0.0202 | 1.4 ± 1.6 | 7.7 ± 5.4 | |
| Controls | 0.0012 ± 0.0019 | 0.0193 ± 0.0143 | 1.2 ± 1.5 | 8.2 ± 4.8 | |
| YRI | Total | 0.0022 ± 0.0014 | 0.0111 ± 0.0059 | 3.8 ± 2.0 | 6.6 ± 1.8 |
| LOAD | 0.0024 ± 0.0015 | 0.0078 | 4.0 ± 2.2 | 7 | |
| Controls | 0.0023 ± 0.0014 | 0.0113 ± 0.0060 | 3.8 ± 2.0 | 6.5 ± 1.9 |
Genome-wide autozygosity determined through FROH estimates and number of ROHs across the eight ethnic groups according to case-control and consanguinity status.
ACD, African–Caribbean from Dominican Republic; AJE, Ashkenazi-Jewish Europeans; ECD, European–Caribbean from Dominican Republic; FCN, French–Canadians; FIN, Finnish Europeans; NWE, North-Western Europeans; SEE, South-Eastern Europeans; YRI, African Yoruba.
Consanguineous subjects in ACD, ECD, FIN, and NWE groups showed a statistically significant (ACD, ECD, NWE), or nearly significant (FIN), lower education level compared to the relative outbred population (on average, 1.4 fewer EDU, Supplementary Tables S6–S13). Conversely, YRI consanguineous subjects showed higher education levels compared to the relative outbred population (Supplementary Table S13). No association was detected between FROH and EDU in the outbred sample (N = 8,648, β = −0.059, SE = 0.196, P = 0.760, heterogeneity-Q = 2.1 × 10–7). However, a highly significant heterogeneity was found across seven ethnic groups, with ACD (β = −3.765, P = 0.077), ECD (β = −2.688, P = 0.059), and FIN (β = −4.959, P < 0.00001) showing a significant, or nearly significant, negative correlation between FROH and EDU, while YRI displayed an opposite correlation (β = 0.870, P = 0.079).
No significant differences were found for mean age between inbred and outbred groups. Consanguinity rates reported for the eight analyzed ethnic groups largely fall within the ranges reported in the literature (
The distribution of APOE genotypes and alleles did not show deviations from HWE (Supplementary Table S14). In addition, APOE genotypes and allele counts differed significantly between outbred and consanguineous subjects only for AJE controls, FCN cases and FIN controls (Supplementary Table S14).
Consanguinity Is Associated With an Increased Risk of LOAD
Consanguinity was significantly associated with LOAD (N = 21,481, OR = 1.262, P = 3.6 × 10–4), independently of APOE∗4 (N = 21,468, OR = 1.237, P = 0.002) and EDU (N = 9,257, OR = 1.274, P = 0.020) (Table 3).
TABLE 3
| Full dataset | ||||||||||||||
| MODEL1 | OR | 95% CI | P | q_p-value | I2 | N | ACD | AJE | ECD | FCN | FIN | NWE | SEE | YRI |
| Consanguinity | 1.262 | 1.111–1.435 | 3.58E-04 | 0.200 | 0.285 | 21,481 | + | – | + | + | + | + | + | + |
| Close consanguinity (1C+2 × 1C+AV) | 1.713 | 1.226–2.394 | 0.002 | 0.385 | 0.049 | 19,227 | + | – | + | – | NA | + | + | NA |
| Distant consanguinity (2C) | 1.207 | 1.054–1.383 | 0.007 | 0.295 | 0.171 | 21,284 | + | – | – | + | + | + | + | + |
| MODEL2 | ||||||||||||||
| Consanguinity | 1.237 | 1.081–1.417 | 0.002 | 0.864 | 0.000 | 21,468 | + | + | + | + | + | + | + | + |
| Close consanguinity (1C+2 × 1C+AV) | 1.798 | 1.267–2.552 | 0.001 | 0.681 | 0.000 | 19,212 | + | + | + | + | NA | + | + | NA |
| Distant consanguinity (2C) | 1.172 | 1.015–1.354 | 0.031 | 0.883 | 0.000 | 21,271 | + | + | – | + | + | + | + | + |
| EDU Subset | ||||||||||||||
| MODEL1 | ||||||||||||||
| Consanguinity | 1.423 | 1.180–1.716 | 2.21E-04 | 0.549 | 0.000 | 9,260 | + | + | + | NA | + | + | + | + |
| Close consanguinity (1C + 2 × 1C + AV) | 2.201 | 1.401–3.455 | 6.16E-04 | 0.487 | 0.000 | 7,619 | + | + | + | NA | NA | + | + | NA |
| Distant consanguinity (2C) | 1.312 | 1.074–1.603 | 0.008 | 0.530 | 0.000 | 9,156 | + | + | – | NA | + | + | + | + |
| MODEL2 | ||||||||||||||
| Consanguinity | 1.418 | 1.165-1.725 | 5.04E-04 | 0.792 | 0.000 | 9,252 | + | + | + | NA | + | + | + | + |
| Close consanguinity (1C + 2 × 1C + AV) | 2.417 | 1.516–3.854 | 2.11E-04 | 0.717 | 0.000 | 7,611 | + | + | + | NA | NA | + | + | NA |
| Distant consanguinity (2C) | 1.287 | 1.042–1.588 | 0.019 | 0.729 | 0.000 | 9,147 | + | + | – | NA | + | + | + | + |
| MODEL3 | ||||||||||||||
| Consanguinity | 1.274 | 1.038–1.562 | 0.020 | 0.483 | 0.000 | 9,252 | + | + | – | NA | + | + | + | + |
| Close consanguinity (1C + 2 × 1C + AV) | 2.014 | 1.240–3.271 | 0.005 | 0.612 | 0.000 | 7,611 | + | + | + | NA | NA | + | + | NA |
| Distant consanguinity (2C) | 1.170 | 0.941–1.455 | 0.158 | 0.459 | 0.000 | 9,147 | + | + | – | NA | + | + | + | + |
Consanguinity increases the risk of LOAD.
ACD, African–Caribbean from Dominican Republic; AJE, Ashkenazi-Jewish Europeans; ECD, European–Caribbean from Dominican Republic; FCN, French–Canadians, FIN, Finnish Europeans; NWE, North-Western Europeans; SEE, South-Eastern Europeans; YRI, African Yoruba. 1C, first-cousins offspring; 2C, second-cousins offspring; 2 × 1C, double-first cousins offspring; AV, avuncular offspring; q p-value, Cochran’s heterogeneity statistic’s p-value; i2, heterogeneity index; ±, summary of effect directions; NA, not available. Values reported in italics are statistically significant.
When testing the association of degree of consanguinity with LOAD by separating close (first cousin/double-first cousin/avuncular offspring) from distant (second cousin offspring) consanguinity, it became clear that the association was driven by close consanguinity (close: N = 19,227, OR = 1.713, P = 0.002; distant: N = 21,284, OR = 1.207, P = 0.007, Table 3). The association reported for close and distant consanguinity was independent of APOE∗4 and EDU (Table 3). When considering the analyses carried out in the smaller EDU subset, the inclusion of APOE∗4 does not reduce statistical estimates of the associations (Table 3). Conversely, the inclusion of EDU as a variable slightly decreases all the associations reported in MODEL3 compared to MODEL2, such that the association of distant consanguinity with LOAD in MODEL3 trends in the same direction but is no longer statistically significant (OR = 1.170, P = 0.158, Table 3).
Autozygosity in the Outbred Population Is Associated With an Increased Risk of LOAD
Table 4 reports the results obtained from the meta-analysis of the association of genome-wide autozygosity determined both by FROH estimates and by NROH across the eight ethnic groups. When considering the full dataset, both FROH and number of ROHs are significantly associated with LOAD, independently of APOE∗4 (Table 4). However, when testing the association of FROH and the number of ROHs with LOAD in the subset with information on EDU, the meta-analysis results are not statistically significant for Table 4, likely reflecting a lack of power rather than an effect of education given that MODELS 1 and 2 are also no longer significant with these sample sizes.
TABLE 4
| Full dataset | ||||||||||||||
| MODEL1 | OR | 95% CI | P | q_p-value | i2 | N | ACD | AJE | ECD | FCN | FIN | NWE | SEE | YRI |
| FROH | 1.204 | 1.018–1.424 | 0.030 | 0.319 | 0.142 | 20,237 | – | – | – | + | – | + | + | + |
| NROH | 1.019 | 1.005–1.034 | 0.006 | 0.220 | 0.261 | 20,237 | + | – | – | + | – | + | + | + |
| MODEL2 | ||||||||||||||
| FROH | 1.222 | 1.021–1.462 | 0.029 | 0.245 | 0.232 | 20,225 | – | – | + | + | – | + | + | + |
| NROH | 1.019 | 1.005–1.034 | 0.007 | 0.241 | 0.237 | 20,225 | – | – | + | + | – | + | + | + |
| EDU subset | ||||||||||||||
| MODEL1 | ||||||||||||||
| FROH | 1.141 | 0.871–1.494 | 0.340 | 0.821 | 0.000 | 8,655 | – | + | – | NA | – | + | + | + |
| NROH | 1.009 | 0.988–1.031 | 0.413 | 0.683 | 0.049 | 8,655 | + | + | – | NA | – | + | + | + |
| MODEL2 | ||||||||||||||
| FROH | 1.180 | 0.887–1.570 | 0.256 | 0.712 | 0.000 | 8,647 | – | + | + | NA | – | + | + | + |
| NROH | 1.011 | 0.988–1.034 | 0.348 | 0.561 | 0.000 | 8,647 | – | + | + | NA | – | + | + | + |
| MODEL3 | ||||||||||||||
| FROH | 1.123 | 0.839–1.503 | 0.435 | 0.392 | 0.046 | 8,647 | – | + | – | NA | – | + | + | + |
| NROH | 1.011 | 0.989–1.034 | 0.334 | 0.258 | 0.224 | 8,647 | – | + | – | NA | – | + | + | + |
Autozygosity in the outbred population increases the risk of LOAD.
ACD, African–Caribbean from Dominican Republic; AJE, Ashkenazi-Jewish Europeans; ECD, European–Caribbean from Dominican Republic; FCN, French–Canadians; FIN, Finnish Europeans; NWE, North-Western Europeans; SEE, South-Eastern Europeans; YRI, African Yoruba. NROH, number of ROHs; Q p-value, Cochran’s heterogeneity statistic’s p-value; i2, heterogeneity index; ±, summary of effect directions; NA, not available. Values reported in italics are statistically significant.
LOAD Genome Is Not Enriched in Rare Recessive Damaging Variants
Given the consistent association of LOAD with consanguinity and autozygosity, we leveraged ADSP WES data to establish whether LOAD subjects showed an enrichment of damaging recessive variants compared to the control population. After merging GWAS imputed data from NWE, SEE, AJE, FIN, and FCN groups with ADSP WES data, we determined that 4,969 subjects were overlapping between the two datasets (AJE = 287; FIN = 47; NWE = 4,424; SEE = 211).
Previous studies have shown that long ROHs are enriched for damaging homozygous variants, with the majority having a MAF ≤ 5% (
When testing the burden of RMHV in LOAD vs. controls, no significant association was detected in the outbred (LOAD:15.15 ± 0.10; Controls:15.35 ± 1.14; P = 0.303) or consanguineous (LOAD:23.97 ± 0.99; Controls:24.48 ± 1.67; P = 0.805) group.
Identification of RPH3AL p.A303V (rs117190076) as RMHV Associated With LOAD
Despite the lack of association between the burden of RMHV and LOAD, we decided to leverage WES data to perform a two-stage recessive-GWAS using the 201 consanguineous subjects identified in ADSP as discovery phase, followed by validation in the remaining 10,469 ADSP subjects. To this aim, we applied the RAFT statistic (
TABLE 5
| Discovery Phase | Validation Phase | |||||||||||||||||||
| Gene | rsID | GnomAD_GlobalMAF | Consequence | CADD | AA_LOAD | Aa_LOAD | aa_LOAD | AA_CONTROLS | Aa_CONTROLS | aa_CONTROLS | GRR | P-value | AA_LOAD | Aa_LOAD | aa_LOAD | AA_CONTROLS | Aa_CONTROLS | aa_CONTROLS | GRR | P-value |
| CLEC1B | rs2273987 | 0.0802 | NP_ 057593.3, p.Ser28Phe | 25 | 117 | 21 | 4 | 58 | 1 | 0 | 404 | 2.66 × 10–10 | 4,773 | 895 | 51 | 3,998 | 710 | 38 | 1.100 | 0.451 |
| KNDC1 | rs11101618 | 0.0220 | NP_ 689856.6, p.Ala128Ser | 24 | 134 | 3 | 3 | 58 | 1 | 0 | 305 | 1.05 × 10–7 | 5,385 | 287 | 4 | 4,413 | 281 | 7 | 1.000 | 1.000 |
| SCAPER | rs200719909 | 0.0004 | NP_ 001339938.1, p.Ala280Val | 15 | 140 | 0 | 2 | 59 | 0 | 0 | 2309 | 2.09 × 10–7 | 5,714 | 4 | 0 | 4,736 | 4 | 0 | — | — |
| IGSF22 | rs78892734 | 0.0375 | NP_ 775859.3, p.Ala1008Thr | 21 | 125 | 4 | 2 | 48 | 0 | 0 | 1987 | 2.85 × 10–7 | 4,838 | 306 | 17 | 4,036 | 232 | 10 | 1.400 | 0.181 |
| C16orf74 | rs141454251 | 0.0092 | NP_ 996850.1, p.Gln12His | 23 | 118 | 0 | 2 | 42 | 0 | 0 | 1779 | 3.59 × 10–7 | 4,982 | 100 | 1 | 3,911 | 64 | 1 | 1.000 | 1.000 |
| SGSH | rs9894254 | 0.0528 | NP_ 000190.1, p.Val361Ile | 23 | 113 | 9 | 3 | 37 | 1 | 0 | 142 | 1.12 × 10–6 | 4,695 | 333 | 13 | 3,627 | 278 | 9 | 1.100 | 0.689 |
| RPH3AL | rs117190076 | 0.0534 | NP_ 001177340.1, p.Ala303Val | 18 | 120 | 14 | 4 | 55 | 3 | 0 | 45 | 2.16 × 10–6 | 4,988 | 603 | 36 | 4,138 | 477 | 16 | 1.900 | 0.001 |
Identification of RPH3AL rs117190076 as rare, deleterious minor homozygote variant associated with LOAD.
AA, major homozygote; Aa, heterozygote; aa, minor homozygote; GRR, Genotype Relative Risk.
Putative UniParental Disomy Does Not Associate With LOAD
During the inbreeding determination process, 56 subjects (6 ACD, 42 NWE, 8 SEE) were found to be potential cases of UPD (Figure 4). The presence of UPD did not show a significant association with an increased risk for LOAD in a logistic regression testing the presence of putative isodisomy compared to the rest of ACD, NWE, and SEE outbred populations (OR = 1.561, P = 0.158). The origin of UPD in our subjects is unknown due to the lack of genotype data from their parents. Nine putative UPD subjects had first-degree relatives genotyped and in these nine cases none of the first-degree relatives showed any evidence of consanguinity or shared long ROHs, suggesting the presence of true isodisomy. Moreover, one of the 12 UPD subjects had ADSP WES data, showing a UPD on chromosome 9p (9p-UDP), was homozygote for the somatic JAK2 V617F (rs77375493) mutation. Since the co-occurrence of JAK2 V617F mutation and 9p-UPD is very common in hematological malignancies (
FIGURE 4

Numerous cases of putative isodisomy misidentified as consanguineous subjects. Fifty-six subjects identified as consanguineous by FSuite v1.0.3 (
Discussion
Our results clearly demonstrate the effect of recent consanguinity and outbred autozygosity in increasing the risk of LOAD consistently across the eight ethnic groups analyzed, independently of APOE∗4 and EDU. Several important features separate our work from previous studies looking at the impact of consanguinity and autozygosity on LOAD. First, and most critically, this is the largest such study to date combining 11,196 cases and 10,296 controls across eight ethnically distinct populations. Second, we went beyond standard super-population definitions of ethnicity and determined European sub-ancestry (NWE, AJE, FCN, FIN, NWE, and SEE), since it has previously been established that these ethnicities have different inbreeding rates (
The overall results suggest the existence of inbreeding depression, which is a recognized phenomenon that is common to polygenic traits in all living organisms (
Notably, despite significant differences in consanguinity rates, autozygosity level, mean age, mean EDU, and APOE frequencies, each of the ethnic groups individually showed significant association (or a non-significant trend in the same direction) for the association of close consanguinity or autozygosity with an increased risk of LOAD. Our results in ACD, ECD, and YRI are reassuringly and not surprisingly, in line with the previous studies (
Previous studies carried out on Europeans of British/Irish descent (
We also leveraged WES data from ADSP to determine the contribution of rare recessive variants in LOAD. The lack of association between the global burden of rare recessive variants and LOAD suggests either the involvement of increased homozygosity at common loci or the existence of specific recessive loci driving the association of consanguinity with LOAD. The two-stage recessive-GWAS we carried out using ADSP WES data showed the association of RPH3AL p.A303V (rs117190076) with LOAD. The RPH3AL gene (also known as NOC2), located on 17p13.3 (OMIM∗604881), encodes for the Rabphilin 3A-like (without C2 domains) protein which plays an essential role in endocrine and exocrine cells, ranging from the accumulation of secretory granules of increased size to impairments in the regulated release of their secretory products (
Interestingly, our work highlighted the presence of potential UPD carriers in the population studied, with prevalence estimates of 0.25% in NWE, 0.74% in SEE and 1.51% in ACD, respectively, showing a trend toward a significant association with increased risk of LOAD. Current estimates of UPD in the general population suggests a general prevalence of 0.05% (1 in 2,000 births) (
One important limitation is that the ethnic stratification, especially for the European groups, led to very small samples in terms of the number of inbred subjects for some of the ancestral groups [e.g., FCN, N = 53; FIN, N = 31 (Supplementary Tables S6–S13)]. Nonetheless, the meta-analytic approach used can greatly mitigate potential biases due to the inclusions of small samples, while providing a better sense of the ethnic-related differences in consanguinity prevalence. Similarly, considering the important contribution of somatic genomic events related to aging such as CHIP, the heterogeneous nature of the specimens used (e.g., whole blood vs. post-mortem brain tissue) in SNP-array genotyping across the different cohorts and samples may have led to uncontrolled biases when determining the impact of true ROHs vs. large genomic deletions of somatic origin.
Overall, these results provide substantial evidence that consanguinity increases risk for LOAD. One might anticipate a change in the genetic architecture of LOAD in the coming decades when more recent cohorts, composed of subjects born after the World War II, will be analyzed. Panmixia and larger effective population sizes have resulted in decreasing autozygosity as the chronological age of a population decreases (
Alzheimer’s Disease Neuroimaging Initiative
Data used in preparation of this article were obtained from the Alzheimer’s Disease Neuroimaging Initiative (ADNI) database (adni.loni.usc.edu). As such. the investigators within the ADNI contributed to the design and implementation of ADNI and/or provided data but did not participate in analysis or writing of this report. A complete listing of ADNI investigators can be found at: http://adni.loni.usc.edu/wp-content/uploads/how_to_apply/ADNI_Acknowledgement_List.pdf.
Statements
Data availability statement
The original contributions presented in the study are included in the article/Supplementary Material, further inquiries can be directed to the corresponding author/s.
Ethics statement
This was a re-analysis of de-identified data available from shared data repositories. The study protocol was granted an exemption by the Stanford Institutional Review Board because the analyses were carried out on “de-identified, off-the-shelf” data. The patients/participants provided their written informed consent to participate in this study.
Author contributions
VN: conceptualization, data curation, formal analysis, investigation, methodology, and writing – original draft preparation. MS: formal analysis, validation, and writing – review and editing. RK: formal analysis, resources, writing – review and editing. AA: writing – review and editing. MG: funding acquisition, supervision, and writing – review and editing. All authors contributed to the article and approved the submitted version.
Funding
This study was supported by the following funding sources: MG: NIH grants AG066515 and AG060747, the McKnight Memory and Cognitive Disorders Award, the Iqbal Farrukh & Asad Jamal Fund, the Show Me Charitable Foundation, and the J. W. Bagley Foundation. AA: Medical Research Council (grant number MR/L016311/1).
Acknowledgments
The authors wish to thank Professor Noah A. Rosenberg (Stanford University) for his insightful comments on the manuscript. MS acknowledges financial support by the EPSRC-funded UCL Centre for Doctoral Training in Medical Imaging (EP/L016478/1). AA holds a Medical Research Council eMedLab Medical Bioinformatics Career Development Fellowship. Data collection and sharing for this project was funded by the Alzheimer’s Disease Neuroimaging Initiative (ADNI) (National Institutes of Health Grant U01 AG024904) and DOD ADNI (Department of Defense award number W81XWH-12-2-0012). ADNI is funded by the National Institute on Aging, the National Institute of Biomedical Imaging and Bioengineering and through generous contributions from the following: AbbVie. Alzheimer’s Association; Alzheimer’s Drug Discovery Foundation; Araclon Biotech; BioClinica. Inc.; Biogen; Bristol-Myers Squibb Company; CereSpir. Inc.; Cogstate; Eisai Inc.; Elan Pharmaceuticals. Inc.; Eli Lilly and Company; EuroImmun; F. Hoffmann-La Roche Ltd. and its affiliated company Genentech. Inc.; Fujirebio; GE HealtControlsare; IXICO Ltd.; Janssen Alzheimer Immunotherapy Research & Development. LLC.; Johnson & Johnson Pharmaceutical Research & Development LLC.; Lumosity; Lundbeck; Merck & Co. Inc.; Meso Scale Diagnostics. LLC.; NeuroRx Research; Neurotrack Technologies; Novartis Pharmaceuticals Corporation; Pfizer Inc.; Piramal Imaging; Servier; Takeda Pharmaceutical Company; and Transition Therapeutics. The Canadian Institutes of Health Research is providing funds to support ADNI clinical sites in Canada. Private sector contributions are facilitated by the Foundation for the National Institutes of Health (www.fnih.org). The grantee organization is the Northern California Institute for Research and Education, and the study is coordinated by the Alzheimer’s Therapeutic Research Institute at the University of Southern California. ADNI data are disseminated by the Laboratory for Neuro Imaging at the University of Southern California. Acknowledgment for Alzheimer Disease Genetic Analysis Data Biological samples and Associated Phenotypic Data used in primary data analyses were stored at Principal Investigators’ institutions, and at the National Cell Repository for Alzheimer’s Disease (NCRAD) at Indiana University funded by NIA. Associated phenotypic data used in primary and secondary data analyses were provided by Principal Investigators, the NIA funded Alzheimer’s Disease Centers (ADCs), and the National Alzheimer’s Coordinating Center (NACC) and stored at Principal Investigators’ Institutions, NCRAD, and at the National Institute on Aging Alzheimer’s Disease Data Storage Site (NIAGADS) at the University of Pennsylvania, funded by NIA. Contributors to the Genetic Analysis Data included Principal Investigators on projects that were individually funded by NIA, other NIH Institutes, private U.S. organizations, or foreign governmental or non-governmental organizations. The NACC database is funded by NIA/NIH Grant U01 AG016976. NACC data are contributed by the NIA funded ADCs: P30 AG019610 (PI Eric Reiman, MD), P30 AG013846 (PI Neil Kowall, MD), P50 AG008702 (PI Scott Small, MD), P50 AG025688 (PI Allan Levey, MD, PhD), P50 AG047266 (PI Todd Golde, MD, Ph.D.). P30 AG010133 (PI Andrew Saykin, PsyD), P50 AG005146 (PI Marilyn Albert, Ph.D.), P50 AG005134 (PI Bradley Hyman, MD, Ph.D.), P50 AG016574 (PI Ronald Petersen, MD, Ph.D.), P50 AG005138 (PI Mary Sano, Ph.D.), P30 AG008051 (PI Steven Ferris, Ph.D.), P30 AG013854 (PI M. Marsel Mesulam, MD), P30 AG008017 (PI Jeffrey Kaye, MD), P30 AG010161 (PI David Bennett, MD), P50 AG047366 (PI Victor Henderson, MD, MS), P30 AG010129 (PI Charles DeCarli, MD), P50 AG016573 (PI Frank LaFerla, Ph.D.), P50 AG016570 (PI Marie-Francoise Chesselet, MD, Ph.D.), P50 AG005131 (PI Douglas Galasko, MD), P50 AG023501 (PI Bruce Miller, MD), P30 AG035982 (PI Russell Swerdlow, MD), P30 AG028383 (PI Linda Van Eldik, Ph.D.), P30 AG010124 (PI John Trojanowski, MD, Ph.D.), P50 AG005133 (PI Oscar Lopez, MD), P50 AG005142 (PI Helena Chui, MD), P30 AG012300 (PI Roger Rosenberg, MD), P50 AG005136 (PI Thomas Montine, MD, Ph.D.), P50 AG033514 (PI Sanjay Asthana, MD, FRCP), P50 AG005681 (PI John Morris, MD), and P50 AG047270 (PI Stephen Strittmatter, MD, Ph.D.).” The genotypic and associated phenotypic data used in the study. “Multi-Site Collaborative Study for Genotype-Phenotype Associations in Alzheimer’s Disease (GenADA)” were provided by the GlaxoSmithKline, R&D Limited. The “Genome Wide Association study of Yoruba in Nigeria” is part of the “Indianapolis – Ibadan Dementia Project (R01 AG009956)” funded by the Division of Neuroscience, National Institute on Aging, National Institutes of Health, ROSMAP study data were provided by the Rush Alzheimer’s Disease Center, Rush University Medical Center, Chicago. Data collection was supported through funding by NIA grants P30AG10161, R01AG15819, R01AG17917, R01AG30146, R01AG36836, U01AG32984, U01AG46152, the Illinois Department of Public Health, and the Translational Genomics Research Institute. Acknowledgment for the Columbia University Study of Caribbean Hispanics with Familial and Sporadic Late Onset Alzheimer’s disease (CIDR): funding support for the “Genetic Consortium for Late Onset Alzheimer’s Disease” was provided through the Division of Neuroscience, NIA. The Genetic Consortium for Late Onset Alzheimer’s Disease includes a genome-wide association study funded as part of the Division of Neuroscience, NIA. Assistance with phenotype harmonization and genotype cleaning, as well as with general study coordination, was provided by Genetic Consortium for Late Onset Alzheimer’s Disease. The AddNeuroMed data are from a public-private partnership supported by EFPIA companies and SMEs as part of InnoMed (Innovative Medicines in Europe), an Integrated Project funded by the European Union of the Sixth Framework program priority FP6-2004-LIFESCIHEALTH-5. Clinical leads responsible for data collection are Iwona Kłoszewska (Lodz), Simon Lovestone (London), Patrizia Mecocci (Perugia), Hilkka Soininen (Kuopio), Magda Tsolaki (Thessaloniki), and Bruno Vellas (Toulouse), imaging leads are Andy Simmons (London), Lars-Olad Wahlund (Stockholm), and Christian Spenger (Zurich) and bioinformatics leads are Richard Dobson (London) and Stephen Newhouse (London).
Conflict of interest
The authors declare that the research was conducted in the absence of any commercial or financial relationships that could be construed as a potential conflict of interest.
Supplementary material
The Supplementary Material for this article can be found online at: https://www.frontiersin.org/articles/10.3389/fgene.2020.629373/full#supplementary-material
References
1
1000 Genomes Project Consortium, AutonA.BrooksL. D.DurbinR. M.GarrisonE. P.KangH. M.et al (2015). A global reference for human genetic variation.Nature52668–74. 10.1038/nature15393
2
AlexanderD. H.NovembreJ.LangeK. (2009). Fast model-based estimation of ancestry in unrelated individuals.Genome Res.191655–1664. 10.1101/gr.094052.109
3
AndrewsS. J.Fulton-HowardB.O’ReillyP.MarcoraE.GoateA. M.Collaborators of the Alzheimer’s Disease Genetics Consortium (2021). Causal associations between modifiable risk factors and the Alzheimer’s Phenome.Ann. Neurol.8954–65. 10.1002/ana.25918
4
BeechamG. W.BisJ. C.MartinE. R.ChoiS. H.DeStefanoA. L.van DuijnC. M.et al (2017). The Alzheimer’s disease sequencing project: study design and sample selection.Neurol. Genet.3:e194.
5
BereczkiE.FrancisP. T.HowlettD.PereiraJ. B.HöglundK.BogstedtA.et al (2016). Synaptic proteins predict cognitive decline in Alzheimer’s disease and lewy body dementia.Alzheimers Dement.121149–1158. 10.1016/j.jalz.2016.04.005
6
BittlesA. H.BlackM. L. (2010). Evolution in health and medicine sackler colloquium: consanguinity. human evolution. and complex diseases.Proc. Natl. Acad. Sci. U.S.A.1071779–1786. 10.1073/pnas.0906079106
7
ChangC. C.ChowC. C.TellierL. C.VattikutiS.PurcellS. M.LeeJ. J. (2015). Second-generation PLINK: rising to the challenge of larger and richer datasets.Gigascience4:7.
8
ChenC. Y.PollackS.HunterD. J.HirshhornJ. N.KraftP.PriceA. L. (2013). Improved ancestry inference using weights from external reference panels.Bioinformatics291399–1406. 10.1093/bioinformatics/btt144
9
ChevietS.WaselleL.RegazziR. (2004). Noc-king out exocrine and endocrine secretion.Trends Cell Biol.14525–528. 10.1016/j.tcb.2004.08.001
10
ClarkD. O.GaoS.LaneK. A.CallahanC. M.BaiyewuO.OgunniyiA.et al (2014). Obesity and 10-year mortality in very old African Americans and Yoruba-Nigerians: exploring the obesity paradox.J. Gerontol. A Biol. Sci. Med. Sci.691162–1169. 10.1093/gerona/glu035
11
CochranJ. N.GeierE. G.BonhamL. W.NewberryJ. S.AmaralM. D.ThompsonM. L.et al (2020). Non-coding and loss-of-function coding variants in TET2 are associated with multiple neurodegenerative diseases.Am. J. Hum. Genet.106632–645. 10.1016/j.ajhg.2020.03.010
12
CraxtonM. (2010). A manual collection of Syt, Esyt, Rph3a, Rph3al, Doc2, and Dblc2 genes from 46 metazoan genomes–an open access resource for neuroscience and evolutionary biology.BMC Genomics11:37. 10.1186/1471-2164-11-37
13
DasS.ForerL.SchönherrS.SidoreC.LockeA. E.KwongA.et al (2016). Next-generation genotype imputation service and methods.Nat. Genet.481284–1287. 10.1038/ng.3656
14
DavidssonP.BogdanovicN.LannfeltL.BlennowK. (2001). Reduced expression of amyloid precursor protein, presenilin-1 and rab3a in cortical brain regions in Alzheimer’s disease.Dement. Geriatr. Cogn. Disord.12243–250. 10.1159/000051266
15
DerbyC. A.KatzM. J.LiptonR. B.HallC. B. (2017). Trends in dementia incidence in a birth cohort analysis of the Einstein aging study.JAMA Neurol.741345–1351. 10.1001/jamaneurol.2017.1964
16
FilippiniN.RaoA.WettenS.GibsonR. A.BorrieM.GuzmanD.et al (2009). Anatomically-distinct genetic associations of APOE epsilon4 allele load with regional cortical atrophy in Alzheimer’s disease.Neuroimage44724–728. 10.1016/j.neuroimage.2008.10.003
17
FukudaM.KannoE.YamamotoA. (2004). Rabphilin and Noc2 are recruited to dense-core vesicles through specific interaction with Rab27A in PC12 cells.J. Biol. Chem.27913065–13075. 10.1074/jbc.m306812200
18
GazalS.SahbatouM.BabronM. C.GéninE.LeuteneggerA. L. (2014). FSuite: exploiting inbreeding in dense SNP chip and exome data.Bioinformatics301940–1941. 10.1093/bioinformatics/btu149
19
GazalS.SahbatouM.BabronM. C.GéninE.LeuteneggerA. L. (2015). High level of inbreeding in final phase of 1000 Genomes Project.Sci. Rep.5:17453.
20
GhaniM.ReitzC.ChengR.VardarajanB. N.JunG.SatoC.et al (2015). Association of long runs of Homozygosity with Alzheimer disease among African American individuals.JAMA Neurol.721313–1323.
21
GhaniM.SatoC.LeeJ. H.ReitzC.MorenoD.MayeuxR.et al (2013). Evidence of recessive Alzheimer disease loci in a Caribbean Hispanic data set: genome-wide survey of runs of homozygosity.JAMA Neurol.701261–1267.
22
GinsbergS. D.MufsonE. J.AlldredM. J.CountsS. E.WuuJ.NixonR. A.et al (2011). Upregulation of select rab GTPases in cholinergic basal forebrain neurons in mild cognitive impairment and Alzheimer’s disease.J. Chem. Neuroanat.42102–110. 10.1016/j.jchemneu.2011.05.012
23
GoldschmidtE.RonenA.RonenI. (1960). Changing marriage systems in the Jewish communities of Israel.Ann. Hum. Genet.24191–204. 10.1111/j.1469-1809.1960.tb01732.x
24
IguchiY.EidL.ParentM.SoucyG.BareilC.RikuY.et al (2016). Exosome secretion is a key pathway for clearance of pathological TDP-43.Brain1393187–3201. 10.1093/brain/aww237
25
JaiswalS.NatarajanP.SilverA. J.GibsonC. J.BickA. G.ShvartzE.et al (2017). Clonal hematopoiesis and risk of atherosclerotic cardiovascular disease.N. Engl. J. Med.377111–121. 10.1056/NEJMoa1701719
26
JakkulaE.RehnströmK.VariloT.PietiläinenO. P.PaunioT.PedersenN. L.et al (2008). The genome-wide patterns of variation expose significant substructure in a founder population.Am. J. Hum. Genet.83787–794. 10.1016/j.ajhg.2008.11.005
27
JoshiP. K.EskoT.MattssonH.EklundN.GandinI.NutileT.et al (2015). Directional dominance on stature and cognition in diverse human populations.Nature523459–462.
28
KangJ. T. L.GoldbergA.EdgeM. D.BeharD. M.RosenbergN. A. (2016). Consanguinity rates predict long runs of Homozygosity in Jewish populations.Hum. Hered.8287–102. 10.1159/000478897
29
KellerM. C.SimonsonM. A.RipkeS.NealeB. M.GejmanP. V.HowriganD. P.et al (2012). Runs of homozygosity implicate autozygosity as a schizophrenia risk factor.PLoS Genet.8:e1002656. 10.1371/journal.pgen.1002656
30
LeeJ. H.ChengR.BarralS.ReitzC.MedranoM.LantiguaR.et al (2011). Identification of novel loci for Alzheimer disease and replication of CLU, PICALM, and BIN1 in caribbean hispanic individuals.Arch. Neurol.68320–328.
31
LiH.WettenS.LiL.St JeanP. L.UpmanyuR.SurhL.et al (2008). Candidate single-nucleotide polymorphisms from a genomewide association study of Alzheimer disease.Arch. Neurol.6545–53.
32
LimE. T.LiuY. P.ChanY.TiinamaijaT.KäräjämäkiA.MadsenE.et al (2014). A novel test for recessive contributions to complex diseases implicates Bardet-Biedl syndrome gene BBS10 in idiopathic type 2 diabetes and obesity.Am. J. Hum. Genet.95509–520. 10.1016/j.ajhg.2014.09.015
33
MägiR.MorrisA. P. (2010). GWAMA: software for genome-wide association meta-analysis.BMC Bioinformatics11:288. 10.1186/1471-2105-11-288
34
MathiasR. A.TaubM. A.GignouxC. R.FuW.MusharoffS.O’ConnorT. D. (2016). A continuum of admixture in the western hemisphere revealed by the African Diaspora genome.Nat. Commun.7:12522.
35
McCarthyS.DasS.KretzschmarW.DelaneauO.WoodA. R.TeumerA.Haplotype Reference Consortium (2016). A reference panel of 64,976 haplotypes for genotype imputation.Nat. Genet.481279–1283. 10.1038/ng.3643
36
McCoyT. H.Jr.CastroV. M.HartK. L.PellegriniA. M.YuS.CaiT.et al (2018). Genome-wide association study of dimensional psychopathology using electronic health records.Biol. Psychiatry831005–1011. 10.1016/j.biopsych.2017.12.004
37
McLarenW.GilL.HuntS. E.RiatH. S.RitchieG. R.ThormannA.et al (2016). The ensembl variant effect predictor.Genome Biol.17:122.
38
NajA. C.JunG.BeechamG. W.WangL. S.VardarajanB. N.BurosJ.et al (2011). Common variants at MS4A4/MS4A6E, CD2AP, CD33 and EPHA1 are associated with late-onset Alzheimer’s disease.Nat. Genet.43436–441.
39
NakkaP.Pattillo SmithS.O’Donnell-LuriaA. H.McManusK. F.23 and Me Research Team, MountainJ. L.et al (2019). Characterization of prevalence and health consequences of uniparental disomy in four million individuals from the general population.Am. J. Hum. Genet.105921–932. 10.1016/j.ajhg.2019.09.016
40
NallsM. A.GuerreiroR. J.Simon-SanchezJ.BrasJ. T.TraynorB. J.GibbsJ. R.et al (2009a). Extended tracts of homozygosity identify novel candidate genes associated with late-onset Alzheimer’s disease.Neurogenetics10183–190. 10.1007/s10048-009-0182-4
41
NallsM. A.Simon-SanchezJ.GibbsJ. R.Paisan-RuizC.BrasJ. T.TanakaT.et al (2009b). Measures of autozygosity in decline: globalization. urbanization. and its implications for medical genetics.PLoS Genet.5:e1000415. 10.1371/journal.pgen.1000415 NODOI
42
NobleJ. M.SchupfN.ManlyJ. J.AndrewsH.TangM. X.MayeuxR. (2017). Secular trends in the incidence of dementia in a Multi-Ethnic community.J. Alzheimers Dis.601065–1075. 10.3233/jad-170300
43
PapenhausenP.SchwartzS.RishegH.KeitgesE.GadiI.BurnsideR. D.et al (2011). UPD detection using homozygosity profiling with a SNP genotyping microarray.Am. J. Med. Genet. A155A757–768. 10.1002/ajmg.a.33939
44
PembertonT. J.AbsherD.FeldmanM. W.MyersR. M.RosenbergN. A.LiJ. Z. (2012). Genomic patterns of homozygosity in worldwide human populations.Am. J. Hum. Genet.91275–292. 10.1016/j.ajhg.2012.06.014
45
PembertonT. J.SzpiechZ. A. (2018). Relationship between deleterious variation, genomic autozygosity, and disease risk: insights from the 1000 Genomes project.Am. J. Hum. Genet.102658–675. 10.1016/j.ajhg.2018.02.013
46
PriceA. L.PattersonN. J.PlengeR. M.WeinblattM. E.ShadickN. A.ReichD. (2006). Principal components analysis corrects for stratification in genome-wide association studies.Nat. Genet.38904–909. 10.1038/ng1847
47
ProitsiP.LuptonM. K.VelayudhanL.NewhouseS.FoghI.TsolakiM.et al (2014). Genetic predisposition to increased blood cholesterol and triglyceride lipid levels and risk of Alzheimer disease: a mendelian randomization analysis.PLoS Med.11:e1001713. 10.1371/journal.pmed.1001713
48
RentzschP.WittenD.CooperG. M.ShendureJ.KircherM. (2019). CADD: predicting the deleteriousness of variants throughout the human genome.Nucleic Acids Res.47D886–D894.
49
Roy-GagnonM. H.MoreauC.BhererC.St-OngeP.SinnettD.LapriseC.et al (2011). Genomic and genealogical investigation of the French Canadian founder population structure.Hum. Genet.129521–531. 10.1007/s00439-010-0945-x
50
RudanI.CampbellH. (2004). Five reasons why inbreeding may have considerable effect on post-reproductive human health.Coll. Antropol.28943–950.
51
RudanI.RudanD.CampbellH.CarothersA.WrightA.Smolej-NarancicN.et al (2003). Inbreeding and risk of late onset complex disease.J. Med. Genet.40925–932. 10.1136/jmg.40.12.925
52
SatizabalC. L.BeiserA. S.ChourakiV.ChêneG.DufouilC.SeshadriS. (2016). Incidence of dementia over three decades in the framingham heart study.N. Engl. J. Med.374523–532. 10.1056/nejmoa1504327
53
SaykinA. J.ShenL.YaoX.KimS.NhoK.RisacherS. L.et al (2015). Genetic studies of quantitative MCI and AD phenotypes in ADNI: progress. opportunities. and plans.Alzheimers Dement.11792–814. 10.1016/j.jalz.2015.05.009
54
SchmandB.SmitJ. H.GeerlingsM. I.LindeboomJ. (1997). The effects of intelligence and education on the development of dementia. A test of the brain reserve hypothesis.Psychol. Med.271337–1344. 10.1017/s0033291797005461
55
SharpE. S.GatzM. (2011). Relationship between education and dementia: an updated systematic review.Alzheimer Dis. Assoc. Disord.25289–304. 10.1097/wad.0b013e318211c83c
56
ShervaR.BaldwinC. T.InzelbergR.VardarajanB.CupplesL. A.LunettaK.et al (2011). Identification of novel candidate genes for Alzheimer’s disease by autozygosity mapping using genome wide SNP data.J. Alzheimers Dis.23349–359. 10.3233/jad-2010-100714
57
SimsR.DwyerS.HaroldD.GerrishA.HollingworthP.ChapmanJ.et al (2011). No evidence that extended tracts of homozygosity are associated with Alzheimer’s disease.Am. J. Med. Genet. B Neuropsychiatr. Genet.156B764–771.
58
TanM. G.LeeC.LeeJ. H.FrancisP. T.WilliamsR. J.RamírezM. J.et al (2014). Decreased rabphilin 3A immunoreactivity in Alzheimer’s disease is associated with Aβ burden.Neurochem. Int.6429–36. 10.1016/j.neuint.2013.10.013
59
VardarajanB. N.SchaidD. J.ReitzC.LantiguaR.MedranoM.Jimènez-VelàzquezI. Z.et al (2015). Inbreeding among caribbean hispanics from the dominican republic and its effects on risk of Alzheimer disease.Genet. Med.17639–643.
60
VélezJ. I.ChandrasekharappaS. C.HenaoE.MartinezA. F.HarperU.JonesM.et al (2013). Pooling/bootstrap-based GWAS (pbGWAS) identifies new loci modifying the age of onset in PSEN1 p.Glu280Ala Alzheimer’s disease.Mol. Psychiatry18568–575.
61
VézinaH.HeyerE.FortierI.OuelletteG.RobitailleY.GauvreauD. (1999). A genealogical study of Alzheimer disease in the Saguenay region of Quebec.Genet. Epidemiol.16412–425.
62
WangL.WheelerD. A.PrchalJ. T. (2016). Acquired uniparental disomy of chromosome 9p in hematologic malignancies.Exp. Hematol.44644–652.
63
ZhangB.GaiteriC.BodeaL. G.WangZ.McElweeJ.PodtelezhnikovA. A.et al (2013). Integrated systems approach identifies genetic nodes and networks in late-onset Alzheimer’s disease.Cell153707–720.
Summary
Keywords
Alzheimer disease, autozygosity, ethnic differences, directional dominance, inbreeding, recessive inheritance, runs of homozygosity (ROH), uniparental isodisomy
Citation
Napolioni V, Scelsi MA, Khan RR, Altmann A and Greicius MD (2021) Recent Consanguinity and Outbred Autozygosity Are Associated With Increased Risk of Late-Onset Alzheimer’s Disease. Front. Genet. 11:629373. doi: 10.3389/fgene.2020.629373
Received
14 November 2020
Accepted
31 December 2020
Published
29 January 2021
Volume
11 - 2020
Edited by
Serena Dato, University of Calabria, Italy
Reviewed by
Giuseppe Passarino, University of Calabria, Italy; Annibale Alessandro Puca, MultiMedica (IRCCS), Italy
Updates

Check for updates
Copyright
© 2021 Napolioni, Scelsi, Khan, Altmann, Greicius and for the Alzheimer’s Disease Neuroimaging Initiative.
This is an open-access article distributed under the terms of the Creative Commons Attribution License (CC BY). The use, distribution or reproduction in other forums is permitted, provided the original author(s) and the copyright owner(s) are credited and that the original publication in this journal is cited, in accordance with accepted academic practice. No use, distribution or reproduction is permitted which does not comply with these terms.
*Correspondence: Valerio Napolioni, valerio.napolioni@unicam.it; napvale@gmail.com
This article was submitted to Genetics of Aging, a section of the journal Frontiers in Genetics
Disclaimer
All claims expressed in this article are solely those of the authors and do not necessarily represent those of their affiliated organizations, or those of the publisher, the editors and the reviewers. Any product that may be evaluated in this article or claim that may be made by its manufacturer is not guaranteed or endorsed by the publisher.