The mutational constraint spectrum quantified from variation in 141,456 humansAbstract Genetic variants that inactivate protein-coding genes are a powerful source of information about the phenotypic consequences of gene disruption: genes that are crucial for the function of an organism will be depleted of such variants in natural populations, whereas non-essential genes will tolerate their accumulation. However, predicted loss-of-function variants are enriched for annotation errors, and tend to be found at extremely low frequencies, so their analysis requires careful variant annotation and very large sample sizes 1 . Here we describe the aggregation of 125,748 exomes and 15,708 genomes from human sequencing studies into the Genome Aggregation Database (gnomAD). We identify 443,769 high-confidence predicted loss-of-function variants in this cohort after filtering for artefacts caused by sequencing and annotation errors. Using an improved model of human mutation rates, we classify human protein-coding genes along a spectrum that represents tolerance to inactivation, validate this classification using data from model organisms and engineered human cells, and show that it can be used to improve the power of gene discovery for both common and rare diseases.
Age-Related Clonal Hematopoiesis Associated with Adverse OutcomesBACKGROUND: The incidence of hematologic cancers increases with age. These cancers are associated with recurrent somatic mutations in specific genes. We hypothesized that such mutations would be detectable in the blood of some persons who are not known to have hematologic disorders. METHODS: We analyzed whole-exome sequencing data from DNA in the peripheral-blood cells of 17,182 persons who were unselected for hematologic phenotypes. We looked for somatic mutations by identifying previously characterized single-nucleotide variants and small insertions or deletions in 160 genes that are recurrently mutated in hematologic cancers. The presence of mutations was analyzed for an association with hematologic phenotypes, survival, and cardiovascular events. RESULTS: Detectable somatic mutations were rare in persons younger than 40 years of age but rose appreciably in frequency with age. Among persons 70 to 79 years of age, 80 to 89 years of age, and 90 to 108 years of age, these clonal mutations were observed in 9.5% (219 of 2300 persons), 11.7% (37 of 317), and 18.4% (19 of 103), respectively. The majority of the variants occurred in three genes: DNMT3A, TET2, and ASXL1. The presence of a somatic mutation was associated with an increase in the risk of hematologic cancer (hazard ratio, 11.1; 95% confidence interval [CI], 3.9 to 32.6), an increase in all-cause mortality (hazard ratio, 1.4; 95% CI, 1.1 to 1.8), and increases in the risks of incident coronary heart disease (hazard ratio, 2.0; 95% CI, 1.2 to 3.4) and ischemic stroke (hazard ratio, 2.6; 95% CI, 1.4 to 4.8). CONCLUSIONS: Age-related clonal hematopoiesis is a common condition that is associated with increases in the risk of hematologic cancer and in all-cause mortality, with the latter possibly due to an increased risk of cardiovascular disease. (Funded by the National Institutes of Health and others.).
A genomic mutational constraint map using variation in 76,156 human genomesA structural variation reference for medical and population geneticsAbstract Structural variants (SVs) rearrange large segments of DNA 1 and can have profound consequences in evolution and human disease 2,3 . As national biobanks, disease-association studies, and clinical genetic testing have grown increasingly reliant on genome sequencing, population references such as the Genome Aggregation Database (gnomAD) 4 have become integral in the interpretation of single-nucleotide variants (SNVs) 5 . However, there are no reference maps of SVs from high-coverage genome sequencing comparable to those for SNVs. Here we present a reference of sequence-resolved SVs constructed from 14,891 genomes across diverse global populations (54% non-European) in gnomAD. We discovered a rich and complex landscape of 433,371 SVs, from which we estimate that SVs are responsible for 25–29% of all rare protein-truncating events per genome. We found strong correlations between natural selection against damaging SNVs and rare SVs that disrupt or duplicate protein-coding sequence, which suggests that genes that are highly intolerant to loss-of-function are also sensitive to increased dosage 6 . We also uncovered modest selection against noncoding SVs in cis -regulatory elements, although selection against protein-truncating SVs was stronger than all noncoding effects. Finally, we identified very large (over one megabase), rare SVs in 3.9% of samples, and estimate that 0.13% of individuals may carry an SV that meets the existing criteria for clinically important incidental findings 7 . This SV resource is freely distributed via the gnomAD browser 8 and will have broad utility in population genetics, disease-association studies, and diagnostic screening.
Functionally significant insulin-like growth factor I receptor mutations in centenariansYousin Suh, Gil Atzmon, Mi-Ook Cho et al.|Proceedings of the National Academy of Sciences|2008 Rather than being a passive, haphazard process of wear and tear, lifespan can be modulated actively by components of the insulin/insulin-like growth factor I (IGFI) pathway in laboratory animals. Complete or partial loss-of-function mutations in genes encoding components of the insulin/IGFI pathway result in extension of life span in yeasts, worms, flies, and mice. This remarkable conservation throughout evolution suggests that altered signaling in this pathway may also influence human lifespan. On the other hand, evolutionary tradeoffs predict that the laboratory findings may not be relevant to human populations, because of the high fitness cost during early life. Here, we studied the biochemical, phenotypic, and genetic variations in a cohort of Ashkenazi Jewish centenarians, their offspring, and offspring-matched controls and demonstrated a gender-specific increase in serum IGFI associated with a smaller stature in female offspring of centenarians. Sequence analysis of the IGF1 and IGF1 receptor (IGF1R) genes of female centenarians showed overrepresentation of heterozygous mutations in the IGF1R gene among centenarians relative to controls that are associated with high serum IGFI levels and reduced activity of the IGFIR as measured in transformed lymphocytes. Thus, genetic alterations in the human IGF1R that result in altered IGF signaling pathway confer an increase in susceptibility to human longevity, suggesting a role of this pathway in modulation of human lifespan.