The ONDRISeq panel: custom-designed next-generation sequencing of genes related to neurodegeneration

Farhan, Sali M K; Dilliott, Allison A; Ghani, Mahdi; Sato, Christine; Liang, Eric; Zhang, Ming; McIntyre, Adam D; Cao, Henian; Racacho, Lemuel; Robinson, John F; Strong, Michael J; Masellis, Mario; St George-Hyslop, Peter; Bulman, Dennis E; Rogaeva, Ekaterina; Hegele, Robert A

doi:10.1038/npjgenmed.2016.32

The ONDRISeq panel: custom-designed next-generation sequencing of genes related to neurodegeneration

Article
Open access
Published: 21 September 2016

Volume 1, article number 16032, (2016)
Cite this article

Download PDF

You have full access to this open access article

npj Genomic Medicine

The ONDRISeq panel: custom-designed next-generation sequencing of genes related to neurodegeneration

Download PDF

Sali M K Farhan^1,2,
Allison A Dilliott^1,2,
Mahdi Ghani³,
Christine Sato³,
Eric Liang¹,
Ming Zhang³,
Adam D McIntyre¹,
Henian Cao¹,
Lemuel Racacho^4,5,
John F Robinson¹,
Michael J Strong^1,6,
Mario Masellis⁷,
Peter St George-Hyslop^3,8,
Dennis E Bulman^4,9,
Ekaterina Rogaeva³ &
Robert A Hegele ORCID: orcid.org/0000-0003-2861-5325^1,2 on behalf of
ONDRI Investigators

4448 Accesses
22 Citations
8 Altmetric
1 Mention
Explore all metrics

Abstract

The Ontario Neurodegenerative Disease Research Initiative (ONDRI) is a multimodal, multi-year, prospective observational cohort study to characterise five diseases: (1) Alzheimer’s disease (AD) or amnestic single or multidomain mild cognitive impairment (aMCI) (AD/MCI); (2) amyotrophic lateral sclerosis (ALS); (3) frontotemporal dementia (FTD); (4) Parkinson’s disease (PD); and (5) vascular cognitive impairment (VCI). The ONDRI Genomics subgroup is investigating the genetic basis of neurodegeneration. We have developed a custom next-generation-sequencing-based panel, ONDRISeq that targets 80 genes known to be associated with neurodegeneration. We processed DNA collected from 216 individuals diagnosed with one of the five diseases, on ONDRISeq. All runs were executed on a MiSeq instrument and subjected to rigorous quality control assessments. We also independently validated a subset of the variant calls using NeuroX (a genome-wide array for neurodegenerative disorders), TaqMan allelic discrimination assay, or Sanger sequencing. ONDRISeq consistently generated high-quality genotyping calls and on average, 92% of targeted bases are covered by at least 30 reads. We also observed 100% concordance for the variants identified via ONDRISeq and validated by other genomic technologies. We were successful in detecting known as well as novel rare variants in 72.2% of cases although not all variants are disease-causing. Using ONDRISeq, we also found that the APOE E4 allele had a frequency of 0.167 in these samples. Our optimised workflow highlights next-generation sequencing as a robust tool in elucidating the genetic basis of neurodegenerative diseases by screening multiple candidate genes simultaneously.

Rare genetic variation in fibronectin 1 (FN1) protects against APOEε4 in Alzheimer’s disease

Article Open access 10 April 2024

An Update on the Adult-Onset Hereditary Cerebellar Ataxias: Novel Genetic Causes and New Diagnostic Approaches

Article Open access 18 May 2024

Analysis and occurrence of biallelic pathogenic repeat expansions in RFC1 in a German cohort of patients with a main clinical phenotype of motor neuron disease

Article Open access 25 June 2024

Introduction

Dementia encompasses a heterogeneous group of neurodegenerative diseases characterised by a progressive decline in cognitive function, language deficiency, and in some cases, motor impairment and behavioural anomalies. Currently, dementia has a global prevalence of 47.5 million cases and an incidence of 7.7 million new cases annually.^1–3 Although today there are no direct treatments available to alter the progressive disease course, early diagnosis has been one of the best predictors of disease outcome.^3,4 Further understanding of the molecular basis of dementia can lead to earlier diagnosis and the eventual development of targeted and efficacious treatment modalities.

Our group is part of the Ontario Neurodegenerative Disease Research Initiative (ONDRI), a multimodal, multi-year, prospective observational cohort study designed to address the effect of small vessel disease in neurodegeneration. ONDRI is recruiting ~600 participants diagnosed with one of the following five diseases: (1) Alzheimer’s disease (AD) or amnestic single- or multidomain mild cognitive impairment (aMCI) (AD/MCI); (2) amyotrophic lateral sclerosis (ALS); (3) frontotemporal dementia (FTD); (4) Parkinson’s disease (PD); and (5) vascular cognitive impairment (VCI).

Genetics is an important risk factor for neurodegenerative disease. Approximately 5–10% of cases with neurodegenerative diseases are familial and can be attributed to several genes.^5–7 However, it is likely we are underestimating the incidence of familial cases based on clinical ascertainment, as the death of presymptomatic individuals may be due to other medical or extrahealth incidents prior to the development of the neurodegenerative syndrome. Furthermore, genetic testing is not universally recommended in the clinical management guidelines of neurodegenerative diseases.^8–13 As such, most neurologists, if they choose to pursue genetic testing, only screen for a small subset of genes and often choose to genotype their patients for highly penetrant and known variants rather than agnostically sequencing all neurodegenerative disease genes. Together, these common clinical ascertainments as well as the high costs associated with genetic testing skew the incidence rates to significantly less than what is perhaps biologically accurate. The five neurodegenerative disorders under study could partly be caused by single, rare, pathogenic variants (monogenic) or multiple, small effect variants acting synergistically to mediate disease expression (oligogenic).

Advancements in next-generation sequencing (NGS) have allowed for efficient genetic variant detection at reduced costs. Currently, there are three main types of NGS applications including: (1) whole-genome sequencing (WGS); (2) whole-exome sequencing (WES); and (3) targeted gene panels.¹⁴ WGS is an indiscriminate approach that evaluates the genetic information in an individual’s entire genome. In contrast, WES targets only the protein-coding regions of the genome as disease-associated variants are significantly over-represented in coding regions.¹⁴ Consequently, WES has been one of the most widely used NGS approaches, however it still presents with several challenges. First, the cost of WES with adequate coverage (i.e., minimum ×30) still remains high at approximately $700. This makes the cumulative cost for studies with a large sample size prohibitively expensive. Second, the amount of genetic variation generated from the exome is excessive and often overwhelming for many researchers and more so for clinicians who may require the patient’s genetic diagnosis to determine whether any genotype-specific treatments are available. Third, WES can generate secondary findings unrelated to the disease of interest, which should be reported to the patient’s primary healthcare provider, in accordance with the guidelines proposed by the American College of Medical Genetics.¹⁵ Thus, in both clinical and research applications, WGS or WES data are still often reduced to focus on likely pathogenic disease-specific loci. In contrast, the use of a targeted gene panel that is clinically focused on the genes underlying the disease(s) of interest, overcomes these issues that often arise when sifting through WGS and WES data.

Herein, we describe the development of a NGS based custom-designed resequencing neurodegeneration gene panel, which we have used to identify genetic variants in neurodegenerative disease cases. ‘ONDRISeq’ allows the screening of patients for variants in 80 genes implicated in neurodegenerative and cerebrovascular disease pathways. However, analysis of 80 genes can still yield an excess of genetic variation. We dichotomised all clinically relevant variants from those of uncertain significance using our integrated custom bioinformatics workflow. Our application of NGS in complex, multifactorial disorders has the potential to identify disease-specific risk markers and potentially, overlapping pathways common across all five diseases.

Results

Study subjects

We recruited 216 participants affected with one of the following disorders: (1) AD/MCI, n=40; (2) ALS, n=22; (3) FTD, n=21; (4) PD, n=56; and (5) VCI, n=77 as part of the ONDRI study (Table 1). The average age of our participants was 69.4±7.8 years. Not surprisingly, individuals diagnosed with ALS were the youngest in our cohort with an average age of 61.9±9.1 years. AD/MCI cases were the oldest patients (mean age of 74.5±6.6 years). The youngest participant in our study is a 40-year-old male diagnosed with ALS; the oldest are four 85-year-old participants (three males, one female); two diagnosed with AD/MCI and two with VCI. In general, sex ratios showed an over representation of males (male:female, 1.8:1.0), which was largely driven by the PD and VCI cases (3.3:1.0 and 2.0:1.0, respectively) similar to the known sex distribution of these disorders in prior population studies. In contrast, in the AD/MCI, ALS, and FTD cohorts, the male:female ratios did not differ considerably (1.5:1.0, 1.2:1.0; and 0.9:1.0, respectively). The self-reported ethnicity of the participants was predominantly Caucasian (82.3%) with some admixture. Overall, participants did not have a family history of neurodegenerative disease and were considered sporadic cases in our study as determined by participant recall, which was confirmed by the participant’s caregiver. Potential confounders such as age, sex, ethnicity and family history did not affect our study objectives or analysis.

Table 1 Patient demographics

Full size table

Quality assessment of ONDRISeq data

In total, 9 independent runs of 24 samples were processed on ONDRISeq (Table 2). All targets across the 216 DNA samples were sufficiently covered (>×30; mean coverage ×76±18; Table 2). On average, 22.8 million of 29.8 million reads passed quality filter equating to 77%. With the exception of the poorest performance run, all ONDRISeq runs had reads passed quality filter of >80%. Overall, 92.7% of all reads were mapped with 95% and 78% of reads mapped in the best and poorest performing runs, respectively. All other ONDRISeq runs had >90% of reads mapped. Of the mapped reads, 87.1% had a Phred quality score of >30 representing a base call accuracy of 99.9%. Similarly, with the exception of the poorest performing run, all ONDRISeq runs had >85% of reads with scores >Q30. Although the poorest performing run produced lower quality data compared with the other 8 ONDRISeq runs, 84.9% of its targets were covered ⩾×30 and were still analysed in our study.

Table 2 Quality control metrics for sequencing runs on ONDRISeq

Full size table

Furthermore, an additional four DNA samples were extracted from brain tissue of deceased individuals. Post autopsy, sections of the brain from all four individuals were frozen for over a decade. However, we were still able to generate adequate sequence calls. Among the four samples, 96% of reads were mapped and each sample had an average coverage of ×71.

ONDRISeq is concordant with NeuroX, TaqMan allelic discrimination assay, and Sanger sequencing

We used three independent genomic techniques, NeuroX, a genome-wide array for neurodegenerative disorders, TaqMan allelic discrimination assays, and Sanger sequencing to assess the concordance with ONDRISeq in variant detection. The NeuroX array captures known polymorphic variants within the genes represented on ONDRISeq; therefore, we evaluated whether ONDRISeq could detect the same variants as NeuroX. In doing so, we processed 115 DNA samples and ONDRISeq detected all 122 non-synonymous variants initially detected by NeuroX. Furthermore, we assessed rare and common, non-synonymous and synonymous variants called by the two platforms and observed 100% concordance between calls. Of note, there were variants detected by ONDRISeq but not included on the NeuroX array. However, there were no false negatives with ONDRISeq: all variants detected by NeuroX were also detected by ONDRISeq. Furthermore, we used a TaqMan allelic discrimination assay to genotype the same 115 DNA samples for APOE. Similarly, we observed 100% concordance between APOE genotyping calls on ONDRISeq and TaqMan.

To explore the rate of false-positive variant calls by ONDRISeq, we performed an independent concordance study for ~10% (n=20) of randomly selected variants from samples that were called as variants by ONDRISeq using Sanger sequencing. Similar with the results of NeuroX and TaqMan allelic discrimination assay, we observed 100% concordance in variants initially detected by ONDRISeq and validated via Sanger sequencing. Thus, there were no false positives with ONDRISeq: all variants called as variants by ONDRISeq were also called as variants by validation using Sanger sequencing.

Clinical utility of ONDRISeq

All DNA samples were independently screened for a hexanucleotide expansion (G₄C₂) within C9orf72, a type of DNA variation that was not detectable by ONDRISeq or NeuroX. Of the 216 samples, only three (1.4%) carried an expansion within C9orf72, two were diagnosed with ALS and one with FTD (Table 3).

Table 3 Other risk variants identified in a cohort of 216 disease cases

Full size table

In total, we found that only 60 out of 216 samples (27.8%) were free from rare (minor allele frequency (MAF) <1%) potentially deleterious variants (missense, nonsense, frameshift, in frame insertions and/or deletions, splicing) in ONDRISeq genes (Table 4). Of the remaining 156 cases, the AD/MCI and FTD cases had the highest variant rate based on ONDRISeq (>80%), although not necessarily disease causative. In the ALS and PD cases, we identified rare coding variants in 72.7% and 71.4% of individuals, respectively. The VCI disease cohort had the lowest number of variant carriers (65%) although still significantly higher than previous reports.^16,17 Furthermore, we tabulated the number of individuals with one, two, or three or more variants. Overall, 76 (48.7%) of 156 individuals carried one variant; 57 (36.5%) carried two variants; and 23 (14.8%) carried three or more variants (Table 4).

Table 4 Diagnostic yield of ONDRISeq in a cohort of 216 disease cases

Full size table

Among the 156 cases with potentially deleterious variants, we identified a total of 266 non-synonymous, rare variants (Table 5), including 107 (40.2%) within genes known to cause the disease with which the patient has been diagnosed (e.g., variation in an AD gene in an AD patient; Table 6). An additional 159 variants (59.8%) were found in genes that were not previously associated with the respective clinical phenotype of the patient, but within a gene responsible for another disease (e.g., variation in FTD gene in an AD patient). Of the 266 variants, which will be reported on in detail upon completion of the ONDRI study of ~600 patients, 62 (23.3%) were previously reported in HGMD and/or ClinVar; whereas 204 (76.7%) were absent from disease databases (Table 5). The majority of variants not found in disease databases were observed in FTD and PD cases (88.9% and 82.6%, respectively); whereas the majority of variants present in disease databases were observed in ALS and VCI cases (35.7% and 28%, respectively; Table 5). On average, we observed four rare variants (MAF<1%) per individual; and 1 variant per individual that met criteria set by ACMG and was considered here, as candidate variants.¹⁵ More rare variants were observed in individuals of African descent (16 rare variants per individual; 2 variants that met ACMG guidelines, per individual). Individuals of South Asian and Chinese origin on average carried 4.5 and 4 rare variants; and 2.5 and 2 variants meeting ACMG guidelines, respectively. These observations are likely due to ascertainment bias in the databases as they typically contain significantly more individuals of European descent than any other ethnic cohort.

Table 5 Variants identified in a cohort of 216 disease cases as detected by ONDRISeq

Full size table

Table 6 Genes associated with amyotrophic lateral sclerosis, frontotemporal dementia, Alzheimer’s disease, Parkinson’s disease, or vascular cognitive impairment as represented on the ONDRISeq targeted resequencing panel

Full size table

Importantly, ONDRISeq is able to provide genotypes for APOE, which is not available through NeuroX and other arrays. In 216 cases, we did not identify a single case of APOE E2/E2 (Table 3). We identified 26 (12%) individuals who had an APOE E2/E3 genotype and 131 (60.6%) individuals who had an APOE E3/E3 genotype (Table 3). In total, 46 (21.3%) individuals were heterozygous for APOE E4 by possessing either an APOE E2/E4 or APOE E3/E4 genotype; whereas 13 (6.02%) individuals were homozygous for APOE E4 (Table 3). Not surprisingly, of the 13 APOE E4/E4 individuals, 7 (53.8%) were diagnosed with AD (Table 3).

Case report: strong evidence of pathogenicity for APP p.Ala713Thr in AD patient

We provide an example of a single neurodegenerative disease case to demonstrate the clinical utility of ONDRISeq and our complementary bioinformatics workflow.

The patient is a 73-year-old male diagnosed with AD. We identified a heterozygous variant, namely g.11248C>T (c.2137G>A), resulting in a missense variant p.Ala713Thr in APP, a gene known to be associated with familial autosomal dominant AD (Figure 1a).¹⁸ The introduction of a polar amino acid within the beta APP domain (amino-acid residues 675–713) is predicted to affect protein function according to multiple in silico analyses and generated a CADD score of 5.483 (Figure 1a,b). The affected codon is also highly conserved in evolution within the APP protein when aligned to a set of diverged species within the animal kingdom (Figure 1c). The variant is very rare with MAF of 0.006% according to Exome Aggregation Consortium (ExAC) and is absent from the 1000 Genomes database and the National Heart, Lung and Blood Institute Exome Variant Server. Furthermore, the patient is the only carrier of p.Ala713Thr in APP, among the 216 samples in our study. However, the variant has been previously observed in AD cases as it is reported in both HGMD and ClinVar databases and has been previously reported in multiple publications.^19–21 Indeed, the variant had sufficient coverage of ×94, nevertheless, we independently validated the presence of the variant using NeuroX and Sanger sequencing (Figure 1a,d,e). The patient is also homozygous for APOE E3/E3.

Discussion

Herein, we describe a NGS based custom-designed resequencing panel to assess genes related to neurodegenerative diseases and small vessel disease. ONDRISeq is a rapid and economical diagnostic approach that screens 80 neurodegenerative genes in parallel. We have processed a total of 216 samples on ONDRISeq in 9 runs with 24 batched samples and evaluated each run using highly stringent quality assessment criteria. With ONDRISeq, we have consistently generated high-quality data and when coupled with our bioinformatics workflow, we have been able to identify rare genetic variants in >70% of patients diagnosed with one of five diseases: AD/MCI, ALS, FTD, PD, or VCI.

The ONDRISeq calls were highly reliable based on validation by three established genetic techniques: NeuroX, a rapid and economical genome-wide genotyping-based neurodegeneration array, TaqMan allelic discrimination assay, and Sanger sequencing. Although NeuroX is able to genotype >250,000 SNPs, the advantage of ONDRISeq is that it is sequencing-based and is able to detect novel variants.²² This way, we can agnostically screen individuals for any novel or known variants within the 80 neurodegenerative genes. Furthermore, although the TaqMan allelic discrimination assay is a rapid genotyping approach, specific probes have to be designed for all SNPs of interest, becoming ultimately costly and inefficient. Also, unlike Sanger sequencing, ONDRISeq is rapid, efficient, and economical. Following library preparation, we are able to analyse the genetic data for 24 samples in <30 h.

We calculated the cost of sequencing 80 genes using standard Sanger sequencing. The total size of ONDRISeq is 971,388 base pairs, which can be processed via ~1,943 PCR reactions (estimation of 500 base pairs per reaction). Had we processed the sequencing reactions in bulk, the cost per sample for Sanger sequencing would have been $38,860 CND per individual. Using NGS-based approaches like WGS or WES with adequate coverage, the price still remains relatively high at $1,400 and $700 CND, respectively (prices based on The Centre for Applied Genomics, Toronto, ON, Canada; www.tcag.ca). Conversely, through strategic cost management we were able to bring our overall expenditures to a highly competitive price of $340 per sample—a reduction of >99% in cost of Sanger sequencing; a >75% reduction relative to WGS, and >50% reduction relative to WES.

Despite its efficiency and rapidity, there are still some limitations with ONDRISeq. First, it can only capture variants within the selected 80 genes, which prevents the discovery of novel disease loci. However, its custom design allows its genetic content to be altered to include novel genomic regions of interest. Second, ONDRISeq is unable to capture multi-nucleotide repeat expansions in genes, a limitation across all NGS platforms.²³ Many neurological diseases such as Huntington’s disease, myotonic dystrophy, Friedreich’s ataxia, Fragile X syndrome, and a subset of spinocerebellar ataxias arising due to multi-nucleotide repeat expansions cannot be detected with current NGS methodologies.^24,25 More recently, a hexanucleotide (G₄C₂) repeat expansion in C9orf72 has been observed in familial and sporadic ALS and FTD cases, and very rarely in PD cases.^26–28 Since its discovery in 2011, it has been one of the top investigated genes as both a diagnostic marker and a therapeutic target.^26,27 C9orf72 alleles can range from 2 to 20 repeats, which are common in the healthy population and are likely benign; 20 to few hundred repeats, which confer risk; or more than few hundred repeats and are pathogenic.^7,29 As such, we independently examined all individuals in our cohort for the C9orf72 expansion using: (1) an amplicon length PCR analysis and (2) a repeat primed PCR analysis. In doing so, we identified that 1.4% of the participants were carriers of a C9orf72 repeat expansion.

Despite these limitations and the complex heterogeneity in the five neurodegenerative diseases that are being assessed with ONDRISeq, we were able to capture rare variants with a probable, but not certain disease association based on allele frequency in the general population and the predictive score of multiple in silico software in 72.2% of cases. As the aetiology of neurodegenerative diseases is often heterogeneous and multiple factors (e.g., genetics, dietary intake, traumatic brain injury, serious infections or toxin exposure) can confer risk to disease onset, we intend to functionally validate the genetic variants, especially the novel variants, to determine their effect size and contribution to disease. Of particular interest are variants in genes with multiple disease associations as they may provide clues on the potential for development of therapy to treat symptoms common across all five neurodegenerative diseases

Materials and methods

Design of ONDRISeq

Using multiple databases, we catalogued literature of neurodegeneration genetic studies. We surveyed 25 content experts (professors, scientists and clinicians within ONDRI) in molecular genetics of neurodegeneration, and used their consensus opinions to select 80 genes within the human genome that were involved in one or more of the five neurodegenerative disorders under study (Table 6). Most genes were selected based on being implicated in neurodegeneration from human genetic studies; however, some of the genes were added based on pathway analysis. Furthermore, some genes were omitted from the ONDRISeq panel due to technical challenges, such as those involving repetitive sequence regions in the genome. This was the case for GBA gene, which is associated with an increased risk of developing PD³⁰ and will thereby be assessed in separate sequencing experiments. Another gene that was omitted from the panel is C9orf72, which contains a repeat expansion and was therefore assessed with a separate genotyping assay as described in subsequent sections.

We designed a composition for detecting variants in the protein-coding regions of 80 genes summing to 1,649 targets. The 80 genes selected have a total target size of 972,388 base pairs. Using the NGS chemistry Nextera Rapid Custom Capture (Illumina, San Diego, CA, USA), we designed a total of 14,510 target specific probes that are each ~80 base pairs in length. For regions that were difficult to sequence, we incorporated additional probes to ensure sufficient coverage during sequencing thereby producing fewer false discoveries. Chromosome scaffold coordinates were obtained from the University of California Santa Cruz Genome Browser using the February 2009 GRCh37/hg19 genome build and were submitted to the Illumina Online Design Studio (Illumina).

Sample collection and DNA isolation

Blood samples were collected from 216 subjects following appropriate and informed consent in accordance with the Research Ethics Board at Parkwood Hospital (London, Ontario, Canada); London Health Sciences Centre (London, Ontario, Canada); Sunnybrook Health Sciences Centre (Toronto, Ontario, Canada); University Health Network-Toronto Western Hospital (Toronto, Ontario, Canada); St Michael's Hospital (Toronto, Ontario, Canada); Centre for Addiction and Mental Health (Toronto, Ontario, Canada); Baycrest Centre for Geriatric Care (Toronto, Ontario, Canada); Hamilton General Hospital (Hamilton, Ontario, Canada); McMaster (Hamilton, Ontario, Canada); Elizabeth Bruyère Hospital (Ottawa, Ontario, Canada); and The Ottawa Hospital (Ottawa, Ontario, Canada).

All clinical diagnoses were supplied by each patient’s healthcare provider in accordance with the criteria from the general ONDRI protocol³¹. DNA was isolated from 4 to 8 ml of blood collected from every participant using the Gentra Puregene Blood kit (Qiagen, Venlo, The Netherlands) according to the manufacturer’s instructions. DNA quality and concentration were initially measured by NanoDrop-1000 Spectrophotometer (Thermo Fisher Scientific, Waltham, MA, USA) and followed by subsequent serial dilutions to obtain ~5 ng/μl. Qubit 2.0 fluorometer technology (Invitrogen, Carlsbad, CA, USA) was then used to measure lower concentrations of DNA at a higher sensitivity.

Library preparation

Libraries were prepared in house using the Nextera Rapid Custom Capture Enrichment kit in accordance with manufacturer’s instructions. DNA samples were processed in sets of 12. DNA samples were fragmented followed by ligation of Nextera Custom Enrichment Kit-specific adapters, amplified via PCR using unique sample barcodes, equimolar pooled, and hybridised to target probes (two cycles of 18 h each). Samples were then amplified again to ensure specificity and greater DNA yield. A small aliquot of each library was analysed using the Agilent 2100 BioAnalyzer (Agilent Technologies, Palo Alto, CA, USA) to ensure adequate yield. The quantity and quality of the final libraries were measured using the KAPA quantitative PCR library quantification kit (KAPA Biosystems, Woburn, MA, USA) using the ViiA 7 Real-Time PCR System (Thermo Fisher Scientific).

Next generation sequencing

All samples were sequenced on the Illumina MiSeq Personal Genome Sequencer (Illumina) using the MiSeq Reagent Kit v3 in accordance with manufacturer’s instructions. Indexed samples were pooled in equimolar ratios of 500 ng. Once combined, 16 pM of denatured pooled library was loaded on to a standard flow-cell on the Illumina MiSeq Personal Sequencer using 2×150 bp paired-end chemistry. Viral PhiX DNA was added as a positive control to ensure sequencer performance. Sequencing quality control was assessed using multiple parameters in Illumina MiSeq Reporter and visualised either in Illumina BaseSpace or locally using Illumina Sequencing Analysis Viewer.

Variant calling

After demultiplexing and adapter trimming, FASTQ files were aligned to the consensus human genome sequence build GRCh37/hg19 using a customised workflow within CLC Bio Genomics Workbench v6.5 (CLC Bio, Aarhus, Denmark) as previously described.³² Similarly, variant annotation was performed using ANNOVAR as previously described³² with additional databases such as CADD³³, HGMD (release 2015.1.), ClinVar³⁴, ExAC³⁵ and our own in-house databases.

APOE genotyping

Furthermore, using ONDRISeq, in addition to screening all samples for variants within APOE, we genotyped all individuals for the APOE risk alleles rs429358(CT) and rs7412(CT). The combination of both individual alleles determines the APOE genotype and is known to be one of the major genetic risk factors for late onset AD.¹⁸ If there are no deletions at these loci, six potential APOE allele combinations are possible (2 alleles×3 possible genotypes): (1) E2/E2; (2) E3/E2; (3) E4/E2; (4) E3/E3; (5) E4/E3; and (6) E4/E4, the latter of which is associated with up to an 11x increased risk in developing AD.^18,36

Variant classification and prioritisation

In general, we followed the guidelines for the interpretation of sequence variants proposed by the American College of Medical Genetics and Genomics and the Association for Molecular Pathology³⁷. We screened for rare variants, which in our study were considered to be variants with MAF<1% based on 1000 Genomes, NHLBI Exome Sequencing Project, and the ExAC databases. Among rare variants, we investigated whether there were any non-synonymous changes (nucleotide substitutions, insertions or deletions) that resulted in missense, nonsense, splicing or frameshift variation. Variants were also assessed in silico using a compilation of prediction programs: PolyPhen-2, SIFT and CADD. HGMD and ClinVar were also integrated to determine the novelty or recurrence of any genetic variation with a specific disease state. More specifically, we were interested in determining how many variants were previously deposited into disease databases. In our study, variants were marked as clinically relevant if they were rare, resulted in non-synonymous changes, were previously observed in individuals with the same disease state, and had values consistent with ‘disease-causing’ based on prediction outcomes of PolyPhen-2, SIFT and CADD, as recommended by ACMG Standards and Guidelines. Importantly, we grouped the variants according to the categories set forth by the ACMG Standards and Guidelines. Alternatively, variants classified here as variants with uncertain clinical significance, were variants that were not reported in disease databases and were observed in genes that are not typically causative of the disease in which the individual is diagnosed with, as represented on Table 6. For example, a variant in a gene that is associated with ALS in a patient diagnosed with VCI. Finally, all ONDRI samples were compared with each other to resolve whether any variants were observed in multiple individuals with the same disease diagnosis. We used this approach to also determine whether the same variant(s) was present in a large subset of ONDRI samples and therefore, was more likely to be an artifact of sequencing or alignment.

Variant validation

To validate variants detected by ONDRISeq, we used three independent genotyping techniques, namely (1) NeuroX, which is an array of specific genotypes that confer risk to several neurodegenerative disease phenotypes;²² (2) TaqMan allelic discrimination; and (3) Sanger sequencing. We processed 115 samples on the NeuroX and the TaqMan allelic discrimination assay and determined the concordance rate between each assay and ONDRISeq. We also randomly selected ~10% of variants to genotype using Sanger sequencing to determine the occurrence if any, of false positives. To test for true negatives, we used DNA from four individuals who were diagnosed with ALS. These individuals were previously tested for genetic variation within SOD1 with no variants identified. Similarly, we did not identify any variants in SOD1 using ONDRISeq. The NeuroX genotyping, TaqMan allelic discrimination assay, and the prior SOD1 testing of ALS samples was performed independently without knowledge of the variants generated using ONDRISeq. Furthermore, different lab personnel completed each validation step to ensure objectivity when evaluating concordance of results.

Variant validation 1: NeuroX

DNA samples were genotyped on NeuroX exome array (Illumina) according to manufacturer’s instructions. NeuroX data were loaded to GenomeStudio (Illumina) and all markers were clustered using the default Gen Call threshold (0.15); duplicate samples (N=2) revealed identical genotypes for all markers with available genotypes (N=268,399).²² Genotypes were converted to PLINK input files, and allele frequencies were calculated. In total, the 115 samples revealed 71,714 polymorphic autosomal markers including 43,129 exonic and 216 splicing variants; among them 39,390 polymorphisms were non-synonymous, as well as 423 stop-gain and 32 stop-loss variants, according to ANNOVAR analyses.²² Average sample call rate was 99.6%, indicating high genotype quality.

Next, 1,047 polymorphic markers, which included 252 exonic variants (229 nonsynonymous and 1 splicing) within the 80 genes of the ONDRISeq targeted sequencing panel, were further processed by removing all noncoding, synonymous and common variants with MAF>1% in any database of 1000Genomes (1000g2014oct_all), Exome Variant Server (esp6500si_all) and ExAC. Variants overlapping segmental duplications were also excluded to avoid possible genotyping error. The remaining variants were filtered to those predicted to have a potential damaging effect on protein function, according to either PolyPhen-2 or SIFT analyses implemented in ANNOVAR.

Variant validation 2: TaqMan allelic discrimination

APOE SNP genotyping was performed using the TaqMan allelic discrimination assay for 115 samples on the 7900HT Fast Real-Time PCR System (Life Technologies, Foster City, CA, USA), and genotypes were identified using automated software (SDS 2.3; Life Technologies). Two TaqMan assays were used to determine the APOE genotype, namely (1) C_3084793_20 (rs429358: APOE codon 112) and (2) C_904973_10 (rs7412; APOE codon 158).

Variant validation 3: Sanger sequencing

Briefly, genomic DNA from the samples was first amplified via PCR, cleaned and purified, and sequenced at the London Regional Genomics Centre. Electropherograms produced were analysed using Applied Biosystems (ABI) SeqScape Software version 2.6 (Thermo Fischer Scientific, Waltham, MA, USA) with the reference sequence of each gene obtained from NCBI GenBank database.

Variant validation 4: SOD1 testing

Screening for genetic variants in the SOD1 gene was performed by PCR followed by standard Sanger sequencing methods, on DNA from four individuals diagnosed with ALS. These steps were performed in other research laboratories prior to this study. Using ONDRISeq, we sequenced DNA from these four individuals to determine whether there were any SOD1 genetic variants. This step allows us to evaluate any true-/false-negative discoveries.

C9orf72 genotyping

All participants were genotyped for the G₄C₂-expansion in C9orf72 using a two-step method: (1) amplicon length analysis and (2) repeat-primed PCR. Experimental procedures are described elsewhere.²⁸

Statistical analysis

The Student’s t-test was used to determine the significance of the difference among patient characteristics within the different neurodegenerative disease cohorts, where appropriate.

References

World Alzheimer Report 2013. Journey of Caring. an Analysis of Long-Term Care for Dementia: Available at https://www.alz.co.uk/research/WorldAlzheimerReport2013.pdf. Accessed 26 August 2015.
WHO takes up the baton on dementia. Lancet Neurol. 14, 455 (2015).
Robinson, L., Tang, E. & Taylor, J. P. Dementia: timely diagnosis and early intervention. BMJ 350, h3029 (2015).
Article Google Scholar
Hebert, L. E., Weuve, J., Scherr, P. A. & Evans, D. A. Alzheimer disease in the United States (2010-2050) estimated using the 2010 census. Neurology 80, 1778–1783 (2013).
Article Google Scholar
Guerreiro, R., Bras, J., Hardy, J. & Singleton, A. Next generation sequencing techniques in neurological diseases: redefining clinical and molecular associations. Hum. Mol. Genet. 23, R47–R53 (2014).
Article CAS Google Scholar
Renton, A. E., Chio, A. & Traynor, B. J. State of play in amyotrophic lateral sclerosis genetics. Nat. Neurosci. 17, 17–23 (2014).
Article CAS Google Scholar
Rohrer, J. D. et al. C9orf72 expansions in frontotemporal dementia and amyotrophic lateral sclerosis. Lancet Neurol. 14, 291–301 (2015).
Article CAS Google Scholar
Miller, R. G. et al. Practice parameter update: the care of the patient with amyotrophic lateral sclerosis: multidisciplinary care, symptom management, and cognitive/behavioral impairment (an evidence-based review): report of the Quality Standards Subcommittee of the American Academy of Neurology. Neurology 73, 1227–1233 (2009).
Article CAS Google Scholar
Miller, R. G. et al. Practice parameter update: the care of the patient with amyotrophic lateral sclerosis: drug, nutritional, and respiratory therapies (an evidence-based review): report of the Quality Standards Subcommittee of the American Academy of Neurology. Neurology 73, 1218–1226 (2009).
Article CAS Google Scholar
Grimes, D. et al. Canadian guidelines on Parkinson's disease. Can. J. Neurol. Sci. 39, S1–30 (2012).
Article Google Scholar
Goldman, J. S. et al. Genetic counseling and testing for Alzheimer disease: joint practice guidelines of the American College of Medical Genetics and the National Society of Genetic Counselors. Genet. Med. 13, 597–605 (2011).
Article Google Scholar
Strong, M. J. et al. Consensus criteria for the diagnosis of frontotemporal cognitive and behavioural syndromes in amyotrophic lateral sclerosis. Amyotroph Lateral Scler. 10, 131–146 (2009).
Article Google Scholar
Jack, C. R. Jr et al. Introduction to the recommendations from the National Institute on Aging-Alzheimer's Association workgroups on diagnostic guidelines for Alzheimer's disease. Alzheimers Dement. 7, 257–262 (2011).
Article Google Scholar
Farhan, S. M. & Hegele, R. A. Exome sequencing: new insights into lipoprotein disorders. Curr. Cardiol. Rep. 16, 507 (2014).
Article Google Scholar
Green, R. C. et al. ACMG recommendations for reporting of incidental findings in clinical exome and genome sequencing. Genet. Med. 15, 565–574 (2013).
Article CAS Google Scholar
Dwyer, R., Skrobot, O. A., Dwyer, J., Munafo, M. & Kehoe, P. G. Using Alzgene-like approaches to investigate susceptibility genes for vascular cognitive impairment. J. Alzheimers Dis. 34, 145–154 (2013).
Article CAS Google Scholar
Sun, J. H. et al. Genetics of vascular dementia: systematic review and meta-analysis. J. Alzheimers Dis. 46, 611–629 (2015).
Article CAS Google Scholar
Bertram, L., McQueen, M. B., Mullin, K., Blacker, D. & Tanzi, R. E. Systematic meta-analyses of Alzheimer disease genetic association studies: the AlzGene database. Nat. Genet. 39, 17–23 (2007).
Article CAS Google Scholar
Rossi, G. et al. A family with Alzheimer disease and strokes associated with A713T mutation of the APP gene. Neurology 63, 910–912 (2004).
Article CAS Google Scholar
Armstrong, J., Boada, M., Rey, M. J., Vidal, N. & Ferrer, I. Familial Alzheimer disease associated with A713T mutation in APP. Neurosci. Lett. 370, 241–243 (2004).
Article CAS Google Scholar
Bernardi, L. et al. AbetaPP A713T mutation in late onset Alzheimer's disease with cerebrovascular lesions. J. Alzheimers Dis. 17, 383–389 (2009).
Article CAS Google Scholar
Ghani, M. et al. Mutation analysis of patients with neurodegenerative disorders using NeuroX array. Neurobiol. Aging 36, 545.e9–e14 (2015).
Article CAS Google Scholar
Singleton, A. B. Exome sequencing: a transformative technology. Lancet Neurol. 10, 942–946 (2011).
Article CAS Google Scholar
La Spada, A. R., Paulson, H. L. & Fischbeck, K. H. Trinucleotide repeat expansion in neurological disease. Ann. Neurol. 36, 814–822 (1994).
Article CAS Google Scholar
Stevens, J. R., Lahue, E. E., Li, G. M. & Lahue, R. S. Trinucleotide repeat expansions catalyzed by human cell-free extracts. Cell Res. 23, 565–572 (2013).
Article CAS Google Scholar
DeJesus-Hernandez, M. et al. Expanded GGGGCC hexanucleotide repeat in noncoding region of C9ORF72 causes chromosome 9p-linked FTD and ALS. Neuron 72, 245–256 (2011).
Article CAS Google Scholar
Renton, A. E. et al. A hexanucleotide repeat expansion in C9ORF72 is the cause of chromosome 9p21-linked ALS-FTD. Neuron 72, 257–268 (2011).
Article CAS Google Scholar
Xi, Z. et al. Investigation of c9orf72 in 4 neurodegenerative disorders. Arch. Neurol. 69, 1583–1590 (2012).
Article Google Scholar
Xi, Z. et al. Jump from pre-mutation to pathologic expansion in C9orf72. Am. J. Hum. Genet. 96, 962–970 (2015).
Article CAS Google Scholar
Sidransky, E. & Lopez, G. The link between the GBA gene and parkinsonism. Lancet Neurol. 11, 986–998 (2012).
Article CAS Google Scholar
Farhan, S. M. K. et al. The Ontario Neurodegenerative Disease Research Initiative (ONDRI). Can. J. Neurol. Sci (in the press).
Johansen, C. T. et al. LipidSeq: a next-generation clinical resequencing panel for monogenic dyslipidemias. J. Lipid Res. 55, 765–772 (2014).
Article CAS Google Scholar
Kircher, M. et al. A general framework for estimating the relative pathogenicity of human genetic variants. Nat. Genet. 46, 310–315 (2014).
Article CAS Google Scholar
Landrum, M. J. et al. ClinVar: public archive of relationships among sequence variation and human phenotype. Nucleic Acids Res. 42, D980–D985 (2014).
Article CAS Google Scholar
Lek, M. et al. Analysis of protein-coding genetic variation in 60,706 humans. Nature 536, 285–291 (2016).
Article CAS Google Scholar
Ghebranious, N., Ivacic, L., Mallum, J. & Dokken, C. Detection of ApoE E2, E3 and E4 alleles using MALDI-TOF mass spectrometry and the homogeneous mass-extend technology. Nucleic Acids Res. 33, e149 (2005).
Article Google Scholar
Richards, S. et al. Standards and guidelines for the interpretation of sequence variants: a joint consensus recommendation of the American College of Medical Genetics and Genomics and the Association for Molecular Pathology. Genet. Med. 17, 405–424 (2015).
Article Google Scholar

Download references

Acknowledgements

We thank and acknowledge the consent and cooperation of all ONDRI participants. Many thanks to the ONDRI investigators (lead investigator: MJS) and the ONDRI governing committees: executive committee; steering committee; publication committee; recruiting clinicians; assessment platforms leaders; and the ONDRI project management team. For a full list of the ONDRI investigators, please visit: www.ONDRI.ca/people. We thank The Thunder Bay Regional Health Sciences Centre, the University of Ottawa Faculty of Medicine, and the Windsor/Essex County ALS Association. The Temerty Family Foundation provided the major infrastructure matching funds. We are grateful to Professor John Hardy from the University College London Institute of Neurology for providing valuable feedback during experimental design. SMKF is supported by the Canadian Institutes of Health Research Frederick Banting and Charles Best Canada Graduate Scholarship.

Author information

Authors and Affiliations

Robarts Research Institute, Schulich School of Medicine and Dentistry, Western University, London, ON, Canada
Sali M K Farhan, Allison A Dilliott, Eric Liang, Adam D McIntyre, Henian Cao, John F Robinson, Michael J Strong & Robert A Hegele
Department of Biochemistry, Schulich School of Medicine and Dentistry, Western University, London, ON, Canada
Sali M K Farhan, Allison A Dilliott & Robert A Hegele
Tanz Centre for Research in Neurodegenerative Diseases, University of Toronto, Toronto, ON, Canada
Mahdi Ghani, Christine Sato, Ming Zhang, Peter St George-Hyslop & Ekaterina Rogaeva
Department of Microbiology and Immunology, Department of Biochemistry, Faculty of Medicine, University of Ottawa, Ottawa, ON, Canada
Lemuel Racacho & Dennis E Bulman
Children's Hospital of Eastern Ontario Research Institute, Ottawa, ON, Canada
Lemuel Racacho
Department of Clinical Neurological Sciences, Schulich School of Medicine & Dentistry, Western University, London, ON, Canada
Michael J Strong
Department of Medicine (Neurology), Sunnybrook Health Sciences Centre, LC Campbell Cognitive Neurology Research Unit, Hurvitz Brain Science Research Program, Sunnybrook Research Institute, University of Toronto, Toronto, ON, Canada
Mario Masellis
Department of Clinical Neurosciences, Cambridge Institute for Medical Research, University of Cambridge, Cambridge, UK
Peter St George-Hyslop
Department of Pediatrics, Faculty of Medicine, University of Ottawa, Ottawa, ON, Canada
Dennis E Bulman

Authors

Sali M K Farhan
View author publications
You can also search for this author in PubMed Google Scholar
Allison A Dilliott
View author publications
You can also search for this author in PubMed Google Scholar
Mahdi Ghani
View author publications
You can also search for this author in PubMed Google Scholar
Christine Sato
View author publications
You can also search for this author in PubMed Google Scholar
Eric Liang
View author publications
You can also search for this author in PubMed Google Scholar
Ming Zhang
View author publications
You can also search for this author in PubMed Google Scholar
Adam D McIntyre
View author publications
You can also search for this author in PubMed Google Scholar
Henian Cao
View author publications
You can also search for this author in PubMed Google Scholar
Lemuel Racacho
View author publications
You can also search for this author in PubMed Google Scholar
John F Robinson
View author publications
You can also search for this author in PubMed Google Scholar
Michael J Strong
View author publications
You can also search for this author in PubMed Google Scholar
Mario Masellis
View author publications
You can also search for this author in PubMed Google Scholar
Peter St George-Hyslop
View author publications
You can also search for this author in PubMed Google Scholar
Dennis E Bulman
View author publications
You can also search for this author in PubMed Google Scholar
Ekaterina Rogaeva
View author publications
You can also search for this author in PubMed Google Scholar
Robert A Hegele
View author publications
You can also search for this author in PubMed Google Scholar

Consortia

ONDRI Investigators

Contributions

Study design: SMKF, MG, MJS, MM, PSGH, DEB, ER and RAH; experimental procedures: SMKF, AAD, MG, EL, MZ, ADM, HC, LR and JR; data analysis: SMKF, AAD, MG, MZ and LR; manuscript preparation: SMKF, MG, PSGH, DEB, ER and RAH; study lead investigator: RAH; ONDRI lead investigator: MJS.

Corresponding author

Correspondence to Robert A Hegele.

Ethics declarations

Competing interests

The authors declare no conflict of interest.

Rights and permissions

This work is licensed under a Creative Commons Attribution 4.0 International License. The images or other third party material in this article are included in the article’s Creative Commons license, unless indicated otherwise in the credit line; if the material is not included under the Creative Commons license, users will need to obtain permission from the license holder to reproduce the material. To view a copy of this license, visit http://creativecommons.org/licenses/by/4.0/

Reprints and permissions

About this article

Cite this article

Farhan, S., Dilliott, A., Ghani, M. et al. The ONDRISeq panel: custom-designed next-generation sequencing of genes related to neurodegeneration. npj Genomic Med 1, 16032 (2016). https://doi.org/10.1038/npjgenmed.2016.32

Download citation

Received: 09 March 2016
Revised: 01 August 2016
Accepted: 05 August 2016
Published: 21 September 2016
DOI: https://doi.org/10.1038/npjgenmed.2016.32
Springer Nature Limited

This article is cited by

White matter hyperintensities in autopsy-confirmed frontotemporal lobar degeneration and Alzheimer’s disease
- Philippe Desmarais
- Andrew F. Gao
- Mario Masellis
Alzheimer's Research & Therapy (2021)
Contribution of rare variant associations to neurodegenerative disease presentation
- Allison A. Dilliott
- Abdalla Abdelhady
- Robert A. Hegele
npj Genomic Medicine (2021)

Editorial Summary

Genomics: New gene panel to screen dementia

Researchers in Canada have developed a rapid, reliable, and economical tool to screen genes associated with dementia. Dementia afflicts nearly 50 million people globally, with an increase of nearly 8 million cases annually. As part of the Ontario Neurodegenerative Disease Research Initiative, a team led by Robert Hegele of Western University created a custom panel called ONDRISeq which uses next generation sequencing to test 80 genes implicated in five neurodegenerative diseases. They evaluated the new panel by testing DNA samples from 216 neurodegenerative patients and comparing the mutations detected by ONDRISeq and three other genomic tests. ONDRISeq detected more variants than other high-throughput approaches, and all of the mutations detected ONDRISeq were confirmed by Sanger sequencing, which is more time-consuming and expensive. The new panel will help scientists diagnose and research the genes behind dementia.

The ONDRISeq panel: custom-designed next-generation sequencing of genes related to neurodegeneration

Abstract

Similar content being viewed by others

Rare genetic variation in fibronectin 1 (FN1) protects against APOEε4 in Alzheimer’s disease

An Update on the Adult-Onset Hereditary Cerebellar Ataxias: Novel Genetic Causes and New Diagnostic Approaches

Analysis and occurrence of biallelic pathogenic repeat expansions in RFC1 in a German cohort of patients with a main clinical phenotype of motor neuron disease

Introduction

Results

Study subjects

Quality assessment of ONDRISeq data

ONDRISeq is concordant with NeuroX, TaqMan allelic discrimination assay, and Sanger sequencing

Clinical utility of ONDRISeq

Case report: strong evidence of pathogenicity for APP p.Ala713Thr in AD patient

Discussion

Materials and methods

Design of ONDRISeq

Sample collection and DNA isolation

Library preparation

Next generation sequencing

Variant calling

APOE genotyping

Variant classification and prioritisation

Variant validation

Variant validation 1: NeuroX

Variant validation 2: TaqMan allelic discrimination

Variant validation 3: Sanger sequencing

Variant validation 4: SOD1 testing

C9orf72 genotyping

Statistical analysis

References

Acknowledgements

Author information

Authors and Affiliations

Consortia

ONDRI Investigators

Contributions

Corresponding author

Ethics declarations

Competing interests

Rights and permissions

About this article

Cite this article

Share this article

This article is cited by

White matter hyperintensities in autopsy-confirmed frontotemporal lobar degeneration and Alzheimer’s disease

Contribution of rare variant associations to neurodegenerative disease presentation

Search

Navigation