Search PubMed⌕ Search

Biomedical subjects

Eric S Lander

Publications and source records attributed to Eric S Lander.

At least 37 records · Page 2Linked to original sources

The case for selection at CCR5-Delta32.

The C-C chemokine receptor 5, 32 base-pair deletion (CCR5-Delta32) allele confers strong resistance to infection by the AIDS virus HIV. Previous studies have suggested that CCR5-Delta32 arose within the past 1,000 y and rose to its present high frequency (5%-14%) in Europe as a result of strong positive selection, perhaps by such selective agents as the bubonic plague or smallpox during the Middle Ages. This hypothesis was based on several lines of evidence, including the absence of the allele outside of Europe and long-range linkage disequilibrium at the locus. We reevaluated this evidence with the benefit of much denser genetic maps and extensive control data. We find that the pattern of genetic variation at CCR5-Delta32 does not stand out as exceptional relative to other loci across the genome. Moreover using newer genetic maps, we estimated that the CCR5-Delta32 allele is likely to have arisen more than 5,000 y ago. While such results can not rule out the possibility that some selection may have occurred at C-C chemokine receptor 5 (CCR5), they imply that the pattern of genetic variation seen at CCR5-Delta32 is consistent with neutral evolution. More broadly, the results have general implications for the design of future studies to detect the signs of positive selection in the human genome.

Alleles↗

Reduced representation bisulfite sequencing for comparative high-resolution DNA methylation analysis.

We describe a large-scale random approach termed reduced representation bisulfite sequencing (RRBS) for analyzing and comparing genomic methylation patterns. BglII restriction fragments were size-selected to 500-600 bp, equipped with adapters, treated with bisulfite, PCR amplified, cloned and sequenced. We constructed RRBS libraries from murine ES cells and from ES cells lacking DNA methyltransferases Dnmt3a and 3b and with knocked-down (kd) levels of Dnmt1 (Dnmt[1(kd),3a-/-,3b-/-]). Sequencing of 960 RRBS clones from Dnmt[1(kd),3a-/-,3b-/-] cells generated 343 kb of non-redundant bisulfite sequence covering 66212 cytosines in the genome. All but 38 cytosines had been converted to uracil indicating a conversion rate of >99.9%. Of the remaining cytosines 35 were found in CpG and 3 in CpT dinucleotides. Non-CpG methylation was >250-fold reduced compared with wild-type ES cells, consistent with a role for Dnmt3a and/or Dnmt3b in CpA and CpT methylation. Closer inspection revealed neither a consensus sequence around the methylated sites nor evidence for clustering of residual methylation in the genome. Our findings indicate random loss rather than specific maintenance of methylation in Dnmt[1(kd),3a-/-,3b-/-] cells. Near-complete bisulfite conversion and largely unbiased representation of RRBS libraries suggest that random shotgun bisulfite sequencing can be scaled to a genome-wide approach.

Animals↗

Gene set enrichment analysis: a knowledge-based approach for interpreting genome-wide expression profiles.

Although genomewide RNA expression analysis has become a routine tool in biomedical research, extracting biological insight from such information remains a major challenge. Here, we describe a powerful analytical method called Gene Set Enrichment Analysis (GSEA) for interpreting gene expression data. The method derives its power by focusing on gene sets, that is, groups of genes that share common biological function, chromosomal location, or regulation. We demonstrate how GSEA yields insights into several cancer-related data sets, including leukemia and lung cancer. Notably, where single-gene analysis finds little similarity between two independent studies of patient survival in lung cancer, GSEA reveals many biological pathways in common. The GSEA method is embodied in a freely available software package, together with an initial database of 1,325 biologically defined gene sets.

Cell Line, Tumor↗

DNA sequence and analysis of human chromosome 18.

Chromosome 18 appears to have the lowest gene density of any human chromosome and is one of only three chromosomes for which trisomic individuals survive to term. There are also a number of genetic disorders stemming from chromosome 18 trisomy and aneuploidy. Here we report the finished sequence and gene annotation of human chromosome 18, which will allow a better understanding of the normal and disease biology of this chromosome. Despite the low density of protein-coding genes on chromosome 18, we find that the proportion of non-protein-coding sequences evolutionarily conserved among mammals is close to the genome-wide average. Extending this analysis to the entire human genome, we find that the density of conserved non-protein-coding sequences is largely uncorrelated with gene density. This has important implications for the nature and roles of non-protein-coding sequence elements.

Aneuploidy↗

A high-density screen for linkage in multiple sclerosis.

To provide a definitive linkage map for multiple sclerosis, we have genotyped the Illumina BeadArray linkage mapping panel (version 4) in a data set of 730 multiplex families of Northern European descent. After the application of stringent quality thresholds, data from 4,506 markers in 2,692 individuals were included in the analysis. Multipoint nonparametric linkage analysis revealed highly significant linkage in the major histocompatibility complex (MHC) on chromosome 6p21 (maximum LOD score [MLS] 11.66) and suggestive linkage on chromosomes 17q23 (MLS 2.45) and 5q33 (MLS 2.18). This set of markers achieved a mean information extraction of 79.3% across the genome, with a Mendelian inconsistency rate of only 0.002%. Stratification based on carriage of the multiple sclerosis-associated DRB1*1501 allele failed to identify any other region of linkage with genomewide significance. However, ordered-subset analysis suggested that there may be an additional locus on chromosome 19p13 that acts independent of the main MHC locus. These data illustrate the substantial increase in power that can be achieved with use of the latest tools emerging from the Human Genome Project and indicate that future attempts to systematically identify susceptibility genes for multiple sclerosis will have to involve large sample sizes and an association-based methodology.

Australia↗

An initial strategy for the systematic identification of functional elements in the human genome by low-redundancy comparative sequencing.

With the recent completion of a high-quality sequence of the human genome, the challenge is now to understand the functional elements that it encodes. Comparative genomic analysis offers a powerful approach for finding such elements by identifying sequences that have been highly conserved during evolution. Here, we propose an initial strategy for detecting such regions by generating low-redundancy sequence from a collection of 16 eutherian mammals, beyond the 7 for which genome sequence data are already available. We show that such sequence can be accurately aligned to the human genome and used to identify most of the highly conserved regions. Although not a long-term substitute for generating high-quality genomic sequences from many mammalian species, this strategy represents a practical initial approach for rapidly annotating the most evolutionarily conserved sequences in the human genome, providing a key resource for the systematic study of human genome function.

Animals↗

A high-resolution linkage-disequilibrium map of the human major histocompatibility complex and first generation of tag single-nucleotide polymorphisms.

Autoimmune, inflammatory, and infectious diseases present a major burden to human health and are frequently associated with loci in the human major histocompatibility complex (MHC). Here, we report a high-resolution (1.9 kb) linkage-disequilibrium (LD) map of a 4.46-Mb fragment containing the MHC in U.S. pedigrees with northern and western European ancestry collected by the Centre d'Etude du Polymorphisme Humain (CEPH) and the first generation of haplotype tag single-nucleotide polymorphisms (tagSNPs) that provide up to a fivefold increase in genotyping efficiency for all future MHC-linked disease-association studies. The data confirm previously identified recombination hotspots in the class II region and allow the prediction of numerous novel hotspots in the class I and class III regions. The region of longest LD maps outside the classic MHC to the extended class I region spanning the MHC-linked olfactory-receptor gene cluster. The extended haplotype homozygosity analysis for recent positive selection shows that all 14 outlying haplotype variants map to a single extended haplotype, which most commonly bears HLA-DRB1*1501. The SNP data, haplotype blocks, and tagSNPs analysis reported here have been entered into a multidimensional Web-based database (GLOVAR), where they can be accessed and viewed in the context of relevant genome annotation. This LD map allowed us to give coordinates for the extremely variable LD structure underlying the MHC.

Haplotypes↗

Systematic discovery of regulatory motifs in human promoters and 3' UTRs by comparison of several mammals.

Comprehensive identification of all functional elements encoded in the human genome is a fundamental need in biomedical research. Here, we present a comparative analysis of the human, mouse, rat and dog genomes to create a systematic catalogue of common regulatory motifs in promoters and 3' untranslated regions (3' UTRs). The promoter analysis yields 174 candidate motifs, including most previously known transcription-factor binding sites and 105 new motifs. The 3'-UTR analysis yields 106 motifs likely to be involved in post-transcriptional regulation. Nearly one-half are associated with microRNAs (miRNAs), leading to the discovery of many new miRNA genes and their likely target genes. Our results suggest that previous estimates of the number of human miRNA genes were low, and that miRNAs regulate at least 20% of human genes. The overall results provide a systematic view of gene regulation in the human, which will be refined as additional mammalian genomes become available.

3' Untranslated Regions↗

Genomic maps and comparative analysis of histone modifications in human and mouse.

We mapped histone H3 lysine 4 di- and trimethylation and lysine 9/14 acetylation across the nonrepetitive portions of human chromosomes 21 and 22 and compared patterns of lysine 4 dimethylation for several orthologous human and mouse loci. Both chromosomes show punctate sites enriched for modified histones. Sites showing trimethylation correlate with transcription starts, while those showing mainly dimethylation occur elsewhere in the vicinity of active genes. Punctate methylation patterns are also evident at the cytokine and IL-4 receptor loci. The Hox clusters present a strikingly different picture, with broad lysine 4-methylated regions that overlay multiple active genes. We suggest these regions represent active chromatin domains required for the maintenance of Hox gene expression. Methylation patterns at orthologous loci are strongly conserved between human and mouse even though many methylated sites do not show sequence conservation notably higher than background. This suggests that the DNA elements that direct the methylation represent only a small fraction of the region or lie at some distance from the site.

Acetylation↗

Assembly of polymorphic genomes: algorithms and application to Ciona savignyi.

Whole-genome assembly is now used routinely to obtain high-quality draft sequence for the genomes of species with low levels of polymorphism. However, genome assembly remains extremely challenging for highly polymorphic species. The difficulty arises because two divergent haplotypes are sequenced together, making it difficult to distinguish alleles at the same locus from paralogs at different loci. We present here a method for assembling highly polymorphic diploid genomes that involves assembling the two haplotypes separately and then merging them to obtain a reference sequence. Our method was developed to assemble the genome of the sea squirt Ciona savignyi, which was sequenced to a depth of 12.7 x from a single wild individual. By comparing finished clones of the two haplotypes we determined that the sequenced individual had an extremely high heterozygosity rate, averaging 4.6% with significant regional variation and rearrangements at all physical scales. Applied to these data, our method produced a reference assembly covering 157 Mb, with N50 contig and scaffold sizes of 47 kb and 989 kb, respectively. Alignment of ESTs indicates that 88% of loci are present at least once and 81% exactly once in the reference assembly. Our method represented loci in a single copy more reliably and achieved greater contiguity than a conventional whole-genome assembly method.

Algorithms↗

Genomewide search and association studies in a Finnish celiac disease population: Identification of a novel locus and replication of the HLA and CTLA4 loci.

It has been reported that celiac disease (CD) is strongly associated with the HLA-DQ2 alleles DQA1*0501 and DQB1*0201. However, this association only accounts for a portion of the genetic component of CD. Several non-HLA loci and candidate genes that potentially contribute to CD susceptibility have been reported, but have not been confirmed. The aim of this study was to identify loci that contribute to disease susceptibility in a CD population from Finland. We performed a genomewide linkage scan and identified two regions of significant linkage to CD (6p and 2q23-32) and one region of suggestive linkage (10p). We also performed targeted typing and analyses that replicated the associations of the HLA and CTLA4 loci.

Adolescent↗

Genome duplication in the teleost fish Tetraodon nigroviridis reveals the early vertebrate proto-karyotype.

Tetraodon nigroviridis is a freshwater puffer fish with the smallest known vertebrate genome. Here, we report a draft genome sequence with long-range linkage and substantial anchoring to the 21 Tetraodon chromosomes. Genome analysis provides a greatly improved fish gene catalogue, including identifying key genes previously thought to be absent in fish. Comparison with other vertebrates and a urochordate indicates that fish proteins have diverged markedly faster than their mammalian homologues. Comparison with the human genome suggests approximately 900 previously unannotated human genes. Analysis of the Tetraodon and human genomes shows that whole-genome duplication occurred in the teleost fish lineage, subsequent to its divergence from mammals. The analysis also makes it possible to infer the basic structure of the ancestral bony vertebrate genome, which was composed of 12 chromosomes, and to reconstruct much of the evolutionary history of ancient and recent chromosome rearrangements leading to the modern human karyotype.

Animals↗

Mapping quantitative trait loci for anxiety in chromosome substitution strains of mice.

Anxious behavior in the mouse is a complex quantitative phenotype that varies widely among inbred mouse strains. We examined a panel of chromosome substitution strains bearing individual A/J chromosomes in an otherwise C57BL/6J background in open-field and light-dark transition tests. Our results confirmed previous reports of quantitative trait loci (QTL) on chromosomes 1, 4, and 15 and identified novel loci on chromosomes 6 and 17. The studies were replicated in two separate laboratories. Systematic differences in the overall activity level were found between the two facilities, but the presence of the QTL was confirmed in both laboratories. We also identified specific effects on open-field defecation and center avoidance and distinguished them from overall open-field activity.

Animals↗

Transcriptional regulatory code of a eukaryotic genome.

DNA-binding transcriptional regulators interpret the genome's regulatory code by binding to specific sequences to induce or repress gene expression. Comparative genomics has recently been used to identify potential cis-regulatory sequences within the yeast genome on the basis of phylogenetic conservation, but this information alone does not reveal if or when transcriptional regulators occupy these binding sites. We have constructed an initial map of yeast's transcriptional regulatory code by identifying the sequence elements that are bound by regulators under various conditions and that are conserved among Saccharomyces species. The organization of regulatory elements in promoters and the environment-dependent use of these elements by regulators are discussed. We find that environment-specific use of regulatory elements predicts mechanistic models for the function of a large population of yeast's transcriptional regulators.

Base Sequence↗

Transgenic rescue demonstrates involvement of the Ian5 gene in T cell development in the rat.

A single point mutation in a novel immune-associated nucleotide gene 5 (Ian5) coincides with severe T cell lymphopenia in BB rats. We used a transgenic rescue approach in lymphopenic BB-derived congenic F344.lyp/lyp rats to determine whether this mutation is responsible for lymphopenia and to establish the functional importance of this novel gene. A 150-kb P1 artificial chromosome (PAC) transgene harboring a wild-type allele of the rat Ian5 gene restored Ian5 transcript and protein levels, completely rescuing the T cell lymphopenia in the F344.lyp/lyp rats. This successful complementation provides direct functional evidence that the Ian5 gene product is essential for maintaining normal T cell levels. It also demonstrates that transgenic rescue in the rat is a practical and definitive method for revealing the function of a novel gene.

Animals↗

Chromosomes 6 and 13 harbor genes that regulate pubertal timing in mouse chromosome substitution strains.

Variation in the onset of puberty among inbred strains of mice suggests that quantitative trait loci (QTLs) affect neurological and hormonal aspects of sexual maturation. Taking a novel approach toward identifying factors that regulate the hypothalamic-pituitary-gonadal (HPG) axis, we evaluated pubertal timing [as assessed by vaginal opening (VO)] in two inbred strains of mice, A/J and C57BL/6J (B6), and in a panel of chromosome substitution strains (CSSs) generated from A/J and B6 mice. In each CSS, a single chromosome from A/J has been substituted in a homozygous fashion for the corresponding chromosome in B6, partitioning the A/J genome into 22 strains with a common host (B6) background. VO occurred significantly earlier in A/J compared with B6 mice. Although the majority of the CSSs assessed had a timing of VO that was similar to the progenitor B6 strain, CSSs for chromosomes 6 and 13 each displayed significantly earlier time of VO than B6 mice. F1 (B6 x CSS) mice for chromosomes 6 and 13 displayed phenotypes that were intermediate between the CSS and B6 strains, suggesting that the trait was inherited in a codominant manner. These findings demonstrate that chromosomes 6 and 13 harbor QTLs that control the timing of VO. Identification of the responsible genes may reveal factors that regulate the maturation of the HPG axis and determine the timing of puberty.

Animals↗

Enhancing linkage analysis of complex disorders: an evaluation of high-density genotyping.

To explore the potential value of recently developed high-density linkage mapping methods in the analysis of complex disease we have regenotyped five nuclear families first studied in the 1996 UK multiple sclerosis linkage genome screen, using Applied Biosystems high-density microsatellite linkage mapping set, the Illumina BeadArray linkage mapping panel (version 3) and the Affymetrix GeneChip Human Mapping 10K array. We found that genotyping success, information extraction and genotyping accuracy were improved with all systems. These improvements were particularly marked with the SNP-based methods (Illumina and Affymetrix), with little difference between these. The extent of additional information extracted is considerable, indicating that reanalysis of existing multiplex families using these newer systems would substantially increase power.

Chromosome Mapping↗

Erralpha and Gabpa/b specify PGC-1alpha-dependent oxidative phosphorylation gene expression that is altered in diabetic muscle.

Recent studies have shown that genes involved in oxidative phosphorylation (OXPHOS) exhibit reduced expression in skeletal muscle of diabetic and prediabetic humans. Moreover, these changes may be mediated by the transcriptional coactivator peroxisome proliferator-activated receptor gamma coactivator-1alpha (PGC-1alpha). By combining PGC-1alpha-induced genome-wide transcriptional profiles with a computational strategy to detect cis-regulatory motifs, we identified estrogen-related receptor alpha (Erralpha) and GA repeat-binding protein alpha as key transcription factors regulating the OXPHOS pathway. Interestingly, the genes encoding these two transcription factors are themselves PGC-1alpha-inducible and contain variants of both motifs near their promoters. Cellular assays confirmed that Erralpha and GA-binding protein a partner with PGC-1alpha in muscle to form a double-positive-feedback loop that drives the expression of many OXPHOS genes. By using a synthetic inhibitor of Erralpha, we demonstrated its key role in PGC-1alpha-mediated effects on gene regulation and cellular respiration. These results illustrate the dissection of gene regulatory networks in a complex mammalian system, elucidate the mechanism of PGC-1alpha action in the OXPHOS pathway, and suggest that Erralpha agonists may ameliorate insulin-resistance in individuals with type 2 diabetes mellitus.

Animals↗