Search PubMedSearch

Biomedical subjects

Shamil R Sunyaev

Publications and source records attributed to Shamil R Sunyaev.

3 recordsLinked to original sources

Inference of elevated mutation rates and variant effects using 700k exomes.

Genomic sequencing is now widely accessible for genetic diagnostics and is emerging as a component of newborn screening. This technological development generates the need to characterize incoming mutations, create comprehensive datasets of genes causing rare Mendelian disorders, and identify pathogenic variants. Large-scale exome sequencing datasets such as Genome Aggregation Database (gnomAD) have been assembled to help address these challenges. The recent release of gnomAD (v4; n = 730,947) uncovers millions of rare coding variants, many of which have arisen more than once by independent recurrent mutations in the rapidly growing recent human population. Here, we use newly developed theoretical understanding of sampling properties of rare variants to estimate key population genetics parameters of practical importance to human genetics such as demography history, mutation rate, and selection. Solely relying on population data, our method Population Inferred Estimates of Selection (PIES) identifies novel genes with loss-of-function mutational hotspots likely due to selection in spermatogonia. PIES efficiently estimates selection coefficients for heterozygous loss-of-function variants. Combining population genetics inference with variant effect predictors, PIES predicts pathogenic missense mutations and improves variant prioritization for genetic diagnostics and newborn screening.

Journal Article

Joint, multifaceted genomic analysis enables diagnosis of diverse, ultra-rare monogenic presentations.

Genomics for rare disease diagnosis has advanced at a rapid pace due to our ability to perform in-depth analyses on individual patients with ultra-rare diseases. The increasing sizes of ultra-rare disease cohorts internationally newly enables cohort-wide analyses for new discoveries, but well-calibrated statistical genetics approaches for jointly analyzing these patients are still under development. The Undiagnosed Diseases Network (UDN) brings multiple clinical, research and experimental centers under the same umbrella across the United States to facilitate and scale case-based diagnostic analyses. Here, we present the first joint analysis of whole genome sequencing data of UDN patients across the network. We introduce new, well-calibrated statistical methods for prioritizing disease genes with de novo recurrence and compound heterozygosity. We also detect pathways enriched with candidate and known diagnostic genes. Our computational analysis, coupled with a systematic clinical review, recapitulated known diagnoses and revealed new disease associations. We further release a software package, RaMeDiES, enabling automated cross-analysis of deidentified sequenced cohorts for new diagnostic and research discoveries. Gene-level findings and variant-level information across the cohort are available in a public-facing browser ( https://dbmi-bgm.github.io/udn-browser/ ). These results show that case-level diagnostic efforts should be supplemented by a joint genomic analysis across cohorts.

Humans

A genome-wide approach for the discovery of novel repeat expansion disorders in the Undiagnosed Diseases Network cohort.

PURPOSE: The Undiagnosed Diseases Network is a National Institutes of Health funded research study that aims to solve a broad clinical spectrum of challenging rare disease cases. Participants receive care from multiple clinical specialists, who collaborate to perform deep phenotyping and state-of-the-art multiomics analyses. As bioinformatics of short-read sequencing has matured, the discovery of repeat expansion disorders (REDs) is accelerating. REDs comprise approximately 60 characterized disorders, which exhibit a broad spectrum of phenotypes. Thus, a largely unbiased genome-wide approach in a phenotypically diverse sample will add to the diagnostic depth, explore the limits of short-read genome analysis, and establish novel candidate RED loci. METHODS: Here, we present a genome-wide analysis of repeat expansions conducted on 1018 genomes from the Undiagnosed Diseases Network. By leveraging 2 distinct bioinformatics tools, ExpansionHunter Denovo and STRling, we showed that repeat expansions can be accurately detected in short-read genomes. RESULTS: We demonstrated that a genotype-first approach can diagnose atypical cases of known REDs and provide valuable clinical insights. We present clinical details on participants with expansions in ATXN7, DMPK, FMR1, GLS, HTT, RFC1, AFF3, and MARCH6. Importantly, we highlight 2 cases of juvenile Huntington disease that were discovered through our analysis. Finally, we present a list of novel candidate short tandem repeats (TR) that could potentially be pathogenic if expanded. CONCLUSION: Importantly, our approach showcases the bioinformatic advancements in genome analysis for RED detection and highlights its practical applications.

Humans