Your browser doesn't support javascript.
loading
Show: 20 | 50 | 100
Results 1 - 5 de 5
Filter
Add more filters











Database
Language
Publication year range
1.
Genome Res ; 34(5): 796-809, 2024 06 25.
Article in English | MEDLINE | ID: mdl-38749656

ABSTRACT

Underrepresented populations are often excluded from genomic studies owing in part to a lack of resources supporting their analyses. The 1000 Genomes Project (1kGP) and Human Genome Diversity Project (HGDP), which have recently been sequenced to high coverage, are valuable genomic resources because of the global diversity they capture and their open data sharing policies. Here, we harmonized a high-quality set of 4094 whole genomes from 80 populations in the HGDP and 1kGP with data from the Genome Aggregation Database (gnomAD) and identified over 153 million high-quality SNVs, indels, and SVs. We performed a detailed ancestry analysis of this cohort, characterizing population structure and patterns of admixture across populations, analyzing site frequency spectra, and measuring variant counts at global and subcontinental levels. We also show substantial added value from this data set compared with the prior versions of the component resources, typically combined via liftOver and variant intersection; for example, we catalog millions of new genetic variants, mostly rare, compared with previous releases. In addition to unrestricted individual-level public release, we provide detailed tutorials for conducting many of the most common quality-control steps and analyses with these data in a scalable cloud-computing environment and publicly release this new phased joint callset for use as a haplotype resource in phasing and imputation pipelines. This jointly called reference panel will serve as a key resource to support research of diverse ancestry populations.


Subject(s)
Databases, Genetic , Genome, Human , Humans , Human Genome Project , High-Throughput Nucleotide Sequencing/methods , Genetic Variation , Genomics/methods
2.
bioRxiv ; 2024 Feb 28.
Article in English | MEDLINE | ID: mdl-36747613

ABSTRACT

Underrepresented populations are often excluded from genomic studies due in part to a lack of resources supporting their analyses. The 1000 Genomes Project (1kGP) and Human Genome Diversity Project (HGDP), which have recently been sequenced to high coverage, are valuable genomic resources because of the global diversity they capture and their open data sharing policies. Here, we harmonized a high quality set of 4,094 whole genomes from HGDP and 1kGP with data from the Genome Aggregation Database (gnomAD) and identified over 153 million high-quality SNVs, indels, and SVs. We performed a detailed ancestry analysis of this cohort, characterizing population structure and patterns of admixture across populations, analyzing site frequency spectra, and measuring variant counts at global and subcontinental levels. We also demonstrate substantial added value from this dataset compared to the prior versions of the component resources, typically combined via liftover and variant intersection; for example, we catalog millions of new genetic variants, mostly rare, compared to previous releases. In addition to unrestricted individual-level public release, we provide detailed tutorials for conducting many of the most common quality control steps and analyses with these data in a scalable cloud-computing environment and publicly release this new phased joint callset for use as a haplotype resource in phasing and imputation pipelines. This jointly called reference panel will serve as a key resource to support research of diverse ancestry populations.

3.
Am J Hum Genet ; 110(9): 1454-1469, 2023 09 07.
Article in English | MEDLINE | ID: mdl-37595579

ABSTRACT

Short-read genome sequencing (GS) holds the promise of becoming the primary diagnostic approach for the assessment of autism spectrum disorder (ASD) and fetal structural anomalies (FSAs). However, few studies have comprehensively evaluated its performance against current standard-of-care diagnostic tests: karyotype, chromosomal microarray (CMA), and exome sequencing (ES). To assess the clinical utility of GS, we compared its diagnostic yield against these three tests in 1,612 quartet families including an individual with ASD and in 295 prenatal families. Our GS analytic framework identified a diagnostic variant in 7.8% of ASD probands, almost 2-fold more than CMA (4.3%) and 3-fold more than ES (2.7%). However, when we systematically captured copy-number variants (CNVs) from the exome data, the diagnostic yield of ES (7.4%) was brought much closer to, but did not surpass, GS. Similarly, we estimated that GS could achieve an overall diagnostic yield of 46.1% in unselected FSAs, representing a 17.2% increased yield over karyotype, 14.1% over CMA, and 4.1% over ES with CNV calling or 36.1% increase without CNV discovery. Overall, GS provided an added diagnostic yield of 0.4% and 0.8% beyond the combination of all three standard-of-care tests in ASD and FSAs, respectively. This corresponded to nine GS unique diagnostic variants, including sequence variants in exons not captured by ES, structural variants (SVs) inaccessible to existing standard-of-care tests, and SVs where the resolution of GS changed variant classification. Overall, this large-scale evaluation demonstrated that GS significantly outperforms each individual standard-of-care test while also outperforming the combination of all three tests, thus warranting consideration as the first-tier diagnostic approach for the assessment of ASD and FSAs.


Subject(s)
Autism Spectrum Disorder , Female , Pregnancy , Humans , Autism Spectrum Disorder/diagnosis , Autism Spectrum Disorder/genetics , Pregnancy Trimester, First , Ultrasonography, Prenatal , Chromosome Mapping , Exome
4.
Nat Commun ; 10(1): 5462, 2019 11 29.
Article in English | MEDLINE | ID: mdl-31784515

ABSTRACT

Human iPSC-derived kidney organoids have the potential to revolutionize discovery, but assessing their consistency and reproducibility across iPSC lines, and reducing the generation of off-target cells remain an open challenge. Here, we profile four human iPSC lines for a total of 450,118 single cells to show how organoid composition and development are comparable to human fetal and adult kidneys. Although cell classes are largely reproducible across time points, protocols, and replicates, we detect variability in cell proportions between different iPSC lines, largely due to off-target cells. To address this, we analyze organoids transplanted under the mouse kidney capsule and find diminished off-target cells. Our work shows how single cell RNA-seq (scRNA-seq) can score organoids for reproducibility, faithfulness and quality, that kidney organoids derived from different iPSC lines are comparable surrogates for human kidney, and that transplantation enhances their formation by diminishing off-target cells.


Subject(s)
Induced Pluripotent Stem Cells/cytology , Kidney/cytology , Organoids/cytology , Animals , Cell Differentiation , Cell Line , Gene Expression Regulation, Developmental , Humans , Induced Pluripotent Stem Cells/metabolism , Induced Pluripotent Stem Cells/transplantation , Kidney/metabolism , Kidney Transplantation , Mice , Organoids/metabolism , Organoids/transplantation , Reproducibility of Results , Sequence Analysis, RNA , Single-Cell Analysis , Transplantation, Heterologous
5.
Cell ; 178(3): 521-535.e23, 2019 07 25.
Article in English | MEDLINE | ID: mdl-31348885

ABSTRACT

Intracellular accumulation of misfolded proteins causes toxic proteinopathies, diseases without targeted therapies. Mucin 1 kidney disease (MKD) results from a frameshift mutation in the MUC1 gene (MUC1-fs). Here, we show that MKD is a toxic proteinopathy. Intracellular MUC1-fs accumulation activated the ATF6 unfolded protein response (UPR) branch. We identified BRD4780, a small molecule that clears MUC1-fs from patient cells, from kidneys of knockin mice and from patient kidney organoids. MUC1-fs is trapped in TMED9 cargo receptor-containing vesicles of the early secretory pathway. BRD4780 binds TMED9, releases MUC1-fs, and re-routes it for lysosomal degradation, an effect phenocopied by TMED9 deletion. Our findings reveal BRD4780 as a promising lead for the treatment of MKD and other toxic proteinopathies. Generally, we elucidate a novel mechanism for the entrapment of misfolded proteins by cargo receptors and a strategy for their release and anterograde trafficking to the lysosome.


Subject(s)
Benzamides/metabolism , Bridged Bicyclo Compounds/pharmacology , Heptanes/pharmacology , Lysosomes/drug effects , Vesicular Transport Proteins/metabolism , Activating Transcription Factor 6/metabolism , Animals , Benzamides/chemistry , Benzamides/pharmacology , Bridged Bicyclo Compounds/therapeutic use , Epithelial Cells/cytology , Epithelial Cells/metabolism , Female , Frameshift Mutation , Heptanes/therapeutic use , Humans , Imidazoline Receptors/antagonists & inhibitors , Imidazoline Receptors/genetics , Imidazoline Receptors/metabolism , Induced Pluripotent Stem Cells/cytology , Induced Pluripotent Stem Cells/metabolism , Kidney/cytology , Kidney/metabolism , Kidney/pathology , Kidney Diseases/metabolism , Kidney Diseases/pathology , Lysosomes/metabolism , Male , Mice , Mice, Transgenic , Mucin-1/chemistry , Mucin-1/genetics , Mucin-1/metabolism , RNA Interference , RNA, Small Interfering/metabolism , Unfolded Protein Response/drug effects , Vesicular Transport Proteins/chemistry
SELECTION OF CITATIONS
SEARCH DETAIL