Search PubMedSearch

Biomedical subjects

Michael N Weedon

Publications and source records attributed to Michael N Weedon.

4 recordsLinked to original sources

Rare Coding Variants Reveal Distinct Genetic Architectures Across Multidimensional Sleep Phenotypes.

Sleep and circadian traits have been widely studied using common variants, but the contribution of rare coding variation remains unclear. We analyzed rare coding variants in 397,065 whole-exome sequenced UK Biobank participants across 36 sleep phenotypes from self-report, diagnoses, sleep medication use and accelerometry, and meta-analyzed results with 171,536 whole-genome sequenced All of Us participants of diverse ancestries, with replication in the Mass General Brigham Biobank (N = 31,275). We identified 260 genes associated with sleep phenotypes, including novel associations with sleep medication use in 29 genes and 24 out of 29 have not previously been reported with any sleep phenotypes. We observed modest but significant rare variant heritability and strong genetic correlations between sleep medication use, insomnia and fatigue. Temporal gene expression trajectory analyses indicate that genes associated with self-reported sleep traits show constant high prenatal expression, whereas genes linked to sleep medication phenotypes exhibit peak expression in the late prenatal period. These findings highlight distinct biological mechanisms captured by different measurement sources of sleep phenotypes and reveal rare-variant-informed targets for therapeutic discovery.

exome sequencing

Streamlining large-scale genomic data management: Insights from the UK Biobank whole-genome sequencing data.

Biobank-scale whole-genome sequencing (WGS) studies are increasingly pivotal in unraveling the genetic bases of diverse health outcomes. However, managing and analyzing these datasets' sheer volume and complexity presents significant challenges. We highlight the annotated genomic data structure (aGDS) format, substantially reducing the WGS data file size while enabling seamless integration of genomic and functional information for comprehensive WGS analyses. The aGDS format yielded 23 chromosome-specific files for the UK Biobank 500k WGS dataset, occupying only 1.10 tebibytes of storage. We develop the vcf2agds toolkit that streamlines the conversion of WGS data from VCF to aGDS format. Additionally, the STAARpipeline equipped with the aGDS files enabled scalable, comprehensive, and functionally informed WGS analysis, facilitating the detection of common and rare coding and noncoding phenotype-genotype associations. Overall, the vcf2agds toolkit and STAARpipeline provide a streamlined solution that facilitates efficient data management and analysis of biobank-scale WGS data across hundreds of thousands of samples.

Humans

Utility of genome sequencing and group-enrichment to support splice variant interpretation in Marfan syndrome.

PURPOSE: To quantify the impact of noncanonical FBN1 splice site variants in undiagnosed Marfan syndrome (MFS), a connective tissue disorder associated with skeletal abnormalities and familial thoracic aortic aneurysm disease (FTAAD). METHODS: A systematic analysis of ultrarare FBN1 variants was performed using genome sequencing data from the 100,000 Genomes Project. Variants were annotated with SpliceAI and the significance of enrichment among individuals with FTAAD was assessed using Fisher's exact test. Experimental validation used RNA sequencing, reverse transcriptase polymerase chain reaction, minigene constructs, and replication analysis was with data from UK Biobank. RESULTS: Using aggregate data for 78,195 individuals, we identified 13,864 singleton single-nucleotide variants in FBN1 of which 21 were predicted to affect splicing (SpliceAI > 0.5). Incidence of candidate splice variants in individuals recruited with FTAAD (9/703) was significantly elevated compared with that seen in non-FTAAD participants (12/77,492; odds ratio = 84, P = 9.7 × 10-14). Additional analysis uncovered a further 14 families harboring 11 different FBN1 splice variants. A total of 20 candidate splice variants in 23 families were identified, of which 70% lay beyond the ±8 splice regions. RNA testing confirmed the predicted splice aberration in 16 of 20 and for 9 of 20, pseudoexonization was the likely splicing anomaly. CONCLUSION: Our findings indicate that noncanonical splice variants may account for approximately 3% of families with undiagnosed FTAAD, highlighting the importance of incorporating analysis of introns and confirmatory RNA testing into genetic testing for Marfan syndrome.

Humans

Whole-genome sequencing in 333,100 individuals reveals rare non-coding single variant and aggregate associations with height.

The role of rare non-coding variation in complex human phenotypes is still largely unknown. To elucidate the impact of rare variants in regulatory elements, we performed a whole-genome sequencing association analysis for height using 333,100 individuals from three datasets: UK Biobank (N&#x2009;=&#x2009;200,003), TOPMed (N&#x2009;=&#x2009;87,652) and All of Us (N&#x2009;=&#x2009;45,445). We performed rare (&#x2009;<&#x2009;0.1% minor-allele-frequency) single-variant and aggregate testing of non-coding variants in regulatory regions based on proximal-regulatory, intergenic-regulatory and deep-intronic annotation. We observed 29 independent variants associated with height at P&#x2009;<&#x2009;after conditioning on previously reported variants, with effect sizes ranging from -7cm to +4.7&#x2009;cm. We also identified and replicated non-coding aggregate-based associations proximal to HMGA1 containing variants associated with a 5&#x2009;cm taller height and of highly-conserved variants in MIR497HG on chromosome 17. We have developed an approach for identifying non-coding rare variants in regulatory regions with large effects from whole-genome sequencing data associated with complex traits.

Humans