Search PubMedSearch

Biomedical subjects

Jessica M Storer

Publications and source records attributed to Jessica M Storer.

2 recordsLinked to original sources

Nuclear mitochondrial sequences in great ape telomere-to-telomere genomes.

Mitochondrial sequences have integrated into the nuclear genome since the origin of eukaryotes. Recent insertions that retain homology with extant mitochondrial DNA (mtDNA), termed NUMTs, confound mtDNA sequence analysis. Here, we use great ape telomere-to-telomere (T2T) genomes to study NUMTs in bonobo, chimpanzee, human, gorilla, and Bornean and Sumatran orangutans. A phylogeny based on shared and lineage-specific NUMTs accurately recapitulates the great ape species tree topology. NUMTs are enriched at nonfunctional nonrepetitive regions of the nuclear genome and depleted within enhancers and coding sequences, suggesting negative selection. We validate the presence of a 76-kb-long heterozygous NUMT in chimpanzee, which is larger than any other NUMT observed in great apes, and find that dozens of NUMTs on the Pan Y Chromosome expanded together with palindromes. Finally, by analyzing intra-specific variation, we confirm that the vast majority of species-specific NUMTs identified in T2T assemblies are fixed or present at high frequencies in each species. Our study highlights NUMTs as a dynamic evolutionary force contributing to shaping ape genomes and is valuable for characterizing mtDNA in great apes.

Journal Article

A complete diploid human genome benchmark for personalized genomics.

Human genome resequencing typically involves mapping reads to a reference genome to call variants; however, this approach suffers from both technical and reference biases, leaving many duplicated and structurally polymorphic regions of the genome unmapped. Consequently, existing variant benchmarks, generated by the same methods, fail to assess these complex regions. To address this limitation, we present a telomere-to-telomere genome benchmark that achieves near-perfect accuracy (i.e. no detectable errors) across 99.4% of the complete, diploid HG002 genome. This benchmark adds 701.4 Mb of autosomal sequence and both sex chromosomes (216.8 Mb), totaling 15.3% of the genome that was absent from prior benchmarks. We also provide a diploid annotation of genes, transposable elements, segmental duplications, and satellite repeats, including 39,144 protein-coding genes across both haplotypes. To facilitate application of the benchmark, we developed tools for measuring the accuracy of sequencing reads, phased variant call sets, and genome assemblies against a diploid reference. Genome-wide analyses show that state-of-the-art de novo assembly methods resolve 2-7% more sequence and outperform variant calling accuracy by an order of magnitude, yielding just one error per 100 kb across 99.9% of the benchmark regions. Adoption of genome-based benchmarking is expected to accelerate the development of cost-effective methods for complete genome sequencing, expanding the reach of genomic medicine to the entire genome and enabling a new era of personalized genomics.

Journal Article