Search PubMed⌕ Search

Biomedical subjects

Jyoti Shetty

Publications and source records attributed to Jyoti Shetty.

14 recordsLinked to original sources

Accurate somatic small variant discovery for multiple sequencing technologies with DeepSomatic.

Somatic variant detection is an integral part of cancer genomics analysis. While most methods have focused on short-read sequencing, long-read technologies offer potential advantages in repeat mapping and variant phasing. We present DeepSomatic, a deep-learning method for detecting somatic small nucleotide variations and insertions and deletions from both short-read and long-read data. The method has modes for whole-genome and whole-exome sequencing and can run on tumor-normal, tumor-only and formalin-fixed paraffin-embedded samples. To train DeepSomatic and help address the dearth of publicly available training and benchmarking data for somatic variant detection, we generated and make openly available the Cancer Standards Long-read Evaluation (CASTLE) dataset of six matched tumor-normal cell line pairs whole-genome sequenced with Illumina, PacBio HiFi and Oxford Nanopore Technologies, along with benchmark variant sets. Across samples, both cell line and patient-derived, and across short-read and long-read sequencing technologies, DeepSomatic consistently outperforms existing callers.

Humans↗

Intrathecally expanded GZMK+/GZMH+ CD8 T cells targeting EBV antigens may reduce severity of Multiple Sclerosis.

Combining cerebrospinal fluid B cell receptor and T cell receptor repertoire analysis with transcriptional/ flow cytometry cellular profiles in hundreds of deeply-phenotyped people with Multiple Sclerosis (pwMS) and controls, we identified intrathecal expansion of anti-viral, cytotoxic, granzymes H/K (GZMH+/GZMK+) double positive (DP) CD8+ T cells that recognize EBV epitopes in pwMS. DP CD8+ T cells are activated and expanded by, and kill autologous, EBV-infected CSF B cell lines in-vitro. Correlations of surrogate transcriptional profiles with clinical and imaging outcomes infer a beneficial role for EBV-targeting DP CD8+ T cells, as untreated pwMS with proportionally higher DP CD8+ T cells to intrathecal B cells accumulate neurological disability slower. MS therapies also increase ratios of beneficial CD8+ T cell responses to intrathecal B cells, consistent with their ability to inhibit disability progression. This study provides indirect evidence that intrathecal EBV infection participates in disability accumulation in pwMS.

Journal Article↗

Severus detects somatic structural variation and complex rearrangements in cancer genomes using long-read sequencing.

For the detection of somatic structural variation (SV) in cancer genomes, long-read sequencing is advantageous over short-read sequencing with respect to mappability and variant phasing. However, most current long-read SV detection methods are not developed for the analysis of tumor genomes characterized by complex rearrangements and heterogeneity. Here, we present Severus, a breakpoint graph-based algorithm for somatic SV calling from long-read cancer sequencing. Severus works with matching normal samples, supports unbalanced cancer karyotypes, can characterize complex multibreak SV patterns and produces haplotype-specific calls. On a comprehensive multitechnology cell line panel, Severus consistently outperforms other long-read and short-read methods in terms of SV detection F1 score (harmonic mean of the precision and recall). We also illustrate that compared to long-read methods, short-read sequencing systematically misses certain classes of somatic SVs, such as insertions or clustered rearrangements. We apply Severus to several clinical cases of pediatric leukemia/lymphoma, revealing clinically relevant cryptic rearrangements missed by standard genomic panels.

Humans↗

DeepSomatic: Accurate somatic small variant discovery for multiple sequencing technologies.

Somatic variant detection is an integral part of cancer genomics analysis. While most methods have focused on short-read sequencing, long-read technologies now offer potential advantages in terms of repeat mapping and variant phasing. We present DeepSomatic, a deep learning method for detecting somatic SNVs and insertions and deletions (indels) from both short-read and long-read data, with modes for whole-genome and exome sequencing, and able to run on tumor-normal, tumor-only, and with FFPE-prepared samples. To help address the dearth of publicly available training and benchmarking data for somatic variant detection, we generated and make openly available a dataset of five matched tumor-normal cell line pairs sequenced with Illumina, PacBio HiFi, and Oxford Nanopore Technologies, along with benchmark variant sets. Across samples and technologies (short-read and long-read), DeepSomatic consistently outperforms existing callers, particularly for indels.

Journal Article↗

The genome sequence of Trypanosoma cruzi, etiologic agent of Chagas disease.

Whole-genome sequencing of the protozoan pathogen Trypanosoma cruzi revealed that the diploid genome contains a predicted 22,570 proteins encoded by genes, of which 12,570 represent allelic pairs. Over 50% of the genome consists of repeated sequences, such as retrotransposons and genes for large families of surface molecules, which include trans-sialidases, mucins, gp63s, and a large novel family (>1300 copies) of mucin-associated surface protein (MASP) genes. Analyses of the T. cruzi, T. brucei, and Leishmania major (Tritryp) genomes imply differences from other eukaryotes in DNA repair and initiation of replication and reflect their unusual mitochondrial DNA. Although the Tritryp lack several classes of signaling molecules, their kinomes contain a large and diverse set of protein kinases and phosphatases; their size and diversity imply previously unknown interactions and regulatory processes, which may be targets for intervention.

Animals↗

Human, mouse, and rat genome large-scale rearrangements: stability versus speciation.

Using paired-end sequences from bacterial artificial chromosomes, we have constructed high-resolution synteny and rearrangement breakpoint maps among human, mouse, and rat genomes. Among the >300 syntenic blocks identified are segments of over 40 Mb without any detected interspecies rearrangements, as well as regions with frequently broken synteny and extensive rearrangements. As closely related species, mouse and rat share the majority of the breakpoints and often have the same types of rearrangements when compared with the human genome. However, the breakpoints not shared between them indicate that mouse rearrangements are more often interchromosomal, whereas intrachromosomal rearrangements are more prominent in rat. Centromeres may have played a significant role in reorganizing a number of chromosomes in all three species. The comparison of the three species indicates that genome rearrangements follow a path that accommodates a delicate balance between maintaining a basic structure underlying all mammalian species and permitting variations that are necessary for speciation.

Animals↗

Confirmation of linkage and refinement of the RP28 locus for autosomal recessive retinitis pigmentosa on chromosome 2p14-p15 in an Indian family.

PURPOSE: To report the linkage analysis of retinitis pigmentosa (RP) in an Indian family. METHODS: Individuals were examined for symptoms of retinitis pigmentosa and their blood samples were withdrawn for genetic analysis. The disorder was tested for linkage to known 14 adRP and 22 arRP loci using microsatellite markers. RESULTS: Seventeen individuals including seven affecteds participated in the study. All affected individuals had typical RP. The age of onset of the disease ranged from 8-18 years. The disorder in this family segregated either as an autosomal recessive trait with pseudodominance or an autosomal dominant trait. Linkage to an autosomal recessive locus RP28 on chromosome 2p14-p15 was positive with a maximum two-point lod score of 3.96 at theta=0 for D2S380. All affected individuals were homozygous for alleles at D2S2320, D2S2397, D2S380, and D2S136. Recombination events placed the minimum critical region (MCR) for the RP28 gene in a 1.06 cM region between D2S2225 and D2S296. CONCLUSIONS: The present data confirmed linkage of arRP to the RP28 locus in a second Indian family. The RP28 locus was previously mapped to a 16 cM region between D2S1337 and D2S286 in a single Indian family. Haplotype analysis in this family has further narrowed the MCR for the RP28 locus to a 1.06 cM region between D2S2225 and D2S296. Of 15 genes reported in the MCR, 14 genes (KIAA0903, OTX1, MDH1, UGP2, VPS54, PELI1, HSPC159, FLJ20080, TRIP-Br2, SLC1A4, KIAA0582, RAB1A, ACTR2, and SPRED2) are either expressed in the eye or retina. Further study needs to be done to test which of these genes is mutated in patients with RP linked to the RP28 locus.

Adolescent↗

Comparison of the genome of the oral pathogen Treponema denticola with other spirochete genomes.

We present the complete 2,843,201-bp genome sequence of Treponema denticola (ATCC 35405) an oral spirochete associated with periodontal disease. Analysis of the T. denticola genome reveals factors mediating coaggregation, cell signaling, stress protection, and other competitive and cooperative measures, consistent with its pathogenic nature and lifestyle within the mixed-species environment of subgingival dental plaque. Comparisons with previously sequenced spirochete genomes revealed specific factors contributing to differences and similarities in spirochete physiology as well as pathogenic potential. The T. denticola genome is considerably larger in size than the genome of the related syphilis-causing spirochete Treponema pallidum. The differences in gene content appear to be attributable to a combination of three phenomena: genome reduction, lineage-specific expansions, and horizontal gene transfer. Genes lost due to reductive evolution appear to be largely involved in metabolism and transport, whereas some of the genes that have arisen due to lineage-specific expansions are implicated in various pathogenic interactions, and genes acquired via horizontal gene transfer are largely phage-related or of unknown function.

ATP-Binding Cassette Transporters↗

Genome sequence of the Brown Norway rat yields insights into mammalian evolution.

The laboratory rat (Rattus norvegicus) is an indispensable tool in experimental medicine and drug development, having made inestimable contributions to human health. We report here the genome sequence of the Brown Norway (BN) rat strain. The sequence represents a high-quality 'draft' covering over 90% of the genome. The BN rat sequence is the third complete mammalian genome to be deciphered, and three-way comparisons with the human and mouse genomes resolve details of mammalian evolution. This first comprehensive analysis includes genes and proteins and their relation to human disease, repeated sequences, comparative genome-wide studies of mammalian orthologous chromosomal regions and rearrangement breakpoints, reconstruction of ancestral karyotypes and the events leading to existing species, rates of variation, and lineage-specific and lineage-independent evolutionary events such as expansion of gene families, orthology relations and protein evolution.

Animals↗

Genetic analysis of an Indian family with members affected with juvenile-onset primary open-angle glaucoma.

PURPOSE: Glaucoma is the second leading cause of blindness. In India, approximately 1.5 million people are blind due to glaucoma. Mutations in the MYOC gene located at the GLC1A locus on chromosome 1q21-q31 have been found in patients with juvenile-onset primary open-angle glaucoma (J-POAG). The purpose of the present study was to identify the genetic cause of glaucoma in a four-generation Indian family affected with J-POAG. METHODS: Peripheral blood samples were obtained from individuals for genomic DNA isolation. To determine if this family was linked to the GLC1A locus, haplotyping analysis was carried out using microsatellite markers from the GLC1A candidate region. Exon-specific primers from exon 3 of the MYOC gene were used to amplify DNA samples from individuals. Mutation analysis was carried out using PCR-SSCP and DNA sequence analyses. RESULTS: Pedigree analysis suggested that glaucoma in this family segregated as an autosomal dominant trait. Of six patients, five had J-POAG and one had adult-onset POAG (A-POAG). Haplotype analysis suggested linkage of this family to the GLC1A locus. Mutation and sequence analyses showed a novel missense mutation, c.821C > G (p.P274R), in the C-terminal olfactomedin domain coded by exon 3 of the MYOC gene. One patient was found to be homozygous for this mutation with a severe phenotype. CONCLUSIONS: This study reports a novel missense mutation in a four-generation Indian family with all but one member affected with J-POAG. The total number of mutations described so far in the MYOC gene, including the one reported here, is 59 with a clustering of 52 mutations in exon 3.

Adolescent↗

Genetic analysis of a high-level vancomycin-resistant isolate of Staphylococcus aureus.

Vancomycin is usually reserved for treatment of serious infections, including those caused by multidrug-resistant Staphylococcus aureus. A clinical isolate of S. aureus with high-level resistance to vancomycin (minimal inhibitory concentration = 1024 microg/ml) was isolated in June 2002. This isolate harbored a 57.9-kilobase multiresistance conjugative plasmid within which Tn1546 (vanA) was integrated. Additional elements on the plasmid encoded resistance to trimethoprim (dfrA), beta-lactams (blaZ), aminoglycosides (aacA-aphD), and disinfectants (qacC). Genetic analyses suggest that the long-anticipated transfer of vancomycin resistance to a methicillin-resistant S. aureus occurred in vivo by interspecies transfer of Tn1546 from a co-isolate of Enterococcus faecalis.

Anti-Bacterial Agents↗

The genome sequence of the malaria mosquito Anopheles gambiae.

Anopheles gambiae is the principal vector of malaria, a disease that afflicts more than 500 million people and causes more than 1 million deaths each year. Tenfold shotgun sequence coverage was obtained from the PEST strain of A. gambiae and assembled into scaffolds that span 278 million base pairs. A total of 91% of the genome was organized in 303 scaffolds; the largest scaffold was 23.1 million base pairs. There was substantial genetic variation within this strain, and the apparent existence of two haplotypes of approximately equal frequency ("dual haplotypes") in a substantial fraction of the genome likely reflects the outbred nature of the PEST strain. The sequence produced a conservative inference of more than 400,000 single-nucleotide polymorphisms that showed a markedly bimodal density distribution. Analysis of the genome sequence revealed strong evidence for about 14,000 protein-encoding transcripts. Prominent expansions in specific families of proteins likely involved in cell adhesion and immunity were noted. An expressed sequence tag analysis of genes regulated by blood feeding provided insights into the physiological adaptations of a hematophagous insect.

Animals↗

The Brucella suis genome reveals fundamental similarities between animal and plant pathogens and symbionts.

The 3.31-Mb genome sequence of the intracellular pathogen and potential bioterrorism agent, Brucella suis, was determined. Comparison of B. suis with Brucella melitensis has defined a finite set of differences that could be responsible for the differences in virulence and host preference between these organisms, and indicates that phage have played a significant role in their divergence. Analysis of the B. suis genome reveals transport and metabolic capabilities akin to soil/plant-associated bacteria. Extensive gene synteny between B. suis chromosome 1 and the genome of the plant symbiont Mesorhizobium loti emphasizes the similarity between this animal pathogen and plant pathogens and symbionts. A limited repertoire of genes homologous to known bacterial virulence factors were identified.

Alphaproteobacteria↗

A physical map of the mouse genome.

A physical map of a genome is an essential guide for navigation, allowing the location of any gene or other landmark in the chromosomal DNA. We have constructed a physical map of the mouse genome that contains 296 contigs of overlapping bacterial clones and 16,992 unique markers. The mouse contigs were aligned to the human genome sequence on the basis of 51,486 homology matches, thus enabling use of the conserved synteny (correspondence between chromosome blocks) of the two genomes to accelerate construction of the mouse map. The map provides a framework for assembly of whole-genome shotgun sequence data, and a tile path of clones for generation of the reference sequence. Definition of the human-mouse alignment at this level of resolution enables identification of a mouse clone that corresponds to almost any position in the human genome. The human sequence may be used to facilitate construction of other mammalian genome maps using the same strategy.

Animals↗