Search PubMed⌕ Search

Biomedical subjects

D R Bentley

Publications and source records attributed to D R Bentley.

At least 19 recordsLinked to original sources

A map of human genome sequence variation containing 1.42 million single nucleotide polymorphisms.

We describe a map of 1.42 million single nucleotide polymorphisms (SNPs) distributed throughout the human genome, providing an average density on available sequence of one SNP every 1.9 kilobases. These SNPs were primarily discovered by two projects: The SNP Consortium and the analysis of clone overlaps by the International Human Genome Sequencing Consortium. The map integrates all publicly available SNPs with described genes and other genomic features. We estimate that 60,000 SNPs fall within exon (coding and untranslated regions), and 85% of exons are within 5 kb of the nearest SNP. Nucleotide diversity varies greatly across the genome, in a manner broadly consistent with a standard population genetic model of human history. This high-density SNP map provides a public resource for defining haplotype variation across the genome, and should help to identify biomedically important genes for diagnosis and therapy.

Chromosome Mapping↗

A physical map of the human genome.

The human genome is by far the largest genome to be sequenced, and its size and complexity present many challenges for sequence assembly. The International Human Genome Sequencing Consortium constructed a map of the whole genome to enable the selection of clones for sequencing and for the accurate assembly of the genome sequence. Here we report the construction of the whole-genome bacterial artificial chromosome (BAC) map and its integration with previous landmark maps and information from mapping efforts focused on specific chromosomal regions. We also describe the integration of sequence data with the map.

Chromosomes, Artificial, Bacterial↗

The physical maps for sequencing human chromosomes 1, 6, 9, 10, 13, 20 and X.

We constructed maps for eight chromosomes (1, 6, 9, 10, 13, 20, X and (previously) 22), representing one-third of the genome, by building landmark maps, isolating bacterial clones and assembling contigs. By this approach, we could establish the long-range organization of the maps early in the project, and all contig extension, gap closure and problem-solving was simplified by containment within local regions. The maps currently represent more than 94% of the euchromatic (gene-containing) regions of these chromosomes in 176 contigs, and contain 96% of the chromosome-specific markers in the human gene map. By measuring the remaining gaps, we can assess chromosome length and coverage in sequenced clones.

Chromosomes, Human, Pair 1↗

A 6.9-Mb high-resolution BAC/PAC contig of human 4p15.3-p16.1, a candidate region for bipolar affective disorder.

Bipolar affective disorder (BPAD) is a complex disease with a significant genetic component and a population lifetime risk of 1%. Our previous work identified a region of human chromosome 4p that showed significant linkage to BPAD in a large pedigree. Here, we report the construction of an accurate, high-resolution physical map of 6.9 Mb of human chromosome 4p15.3-p16.1, which includes an 11-cM (5.8 Mb) critical region for BPAD. The map consists of 460 PAC and BAC clones ordered by a combination of STS content analysis and restriction fragment fingerprinting, with a single approximately 300-kb gap remaining. A total of 289 new and existing markers from a wide range of sources have been localized on the contig, giving an average marker resolution of 1 marker/23 kb. The STSs include 57 ESTs, 9 of which represent known genes. This contig is an essential preliminary to the identification of candidate genes that predispose to bipolar affective disorder, to the completion of the sequence of the region, and to the development of a high-density SNP map.

Bipolar Disorder↗

Long-range comparison of human and mouse SCL loci: localized regions of sensitivity to restriction endonucleases correspond precisely with peaks of conserved noncoding sequences.

Long-range comparative sequence analysis provides a powerful strategy for identifying conserved regulatory elements. The stem cell leukemia (SCL) gene encodes a bHLH transcription factor with a pivotal role in hemopoiesis and vasculogenesis, and it displays a highly conserved expression pattern. We present here a detailed sequence comparison of 193 kb of the human SCL locus to 234 kb of the mouse SCL locus. Four new genes have been identified together with an ancient mitochondrial insertion in the human locus. The SCL gene is flanked upstream by the SIL gene and downstream by the MAP17 gene in both species, but the gene order is not collinear downstream from MAP17. To facilitate rapid identification of candidate regulatory elements, we have developed a new sequence analysis tool (SynPlot) that automates the graphical display of large-scale sequence alignments. Unlike existing programs, SynPlot can display the locus features of more than one sequence, thereby indicating the position of homology peaks relative to the structure of all sequences in the alignment. In addition, high-resolution analysis of the chromatin structure of the mouse SCL gene permitted the accurate positioning of localized zones accessible to restriction endonucleases. Zones known to be associated with functional regulatory regions were found to correspond precisely with peaks of human/mouse homology, thus demonstrating that long-range human/mouse sequence comparisons allow accurate prediction of the extent of accessible DNA associated with active regulatory regions.

Animals↗

An SNP map of human chromosome 22.

The human genome sequence will provide a reference for measuring DNA sequence variation in human populations. Sequence variants are responsible for the genetic component of individuality, including complex characteristics such as disease susceptibility and drug response. Most sequence variants are single nucleotide polymorphisms (SNPs), where two alternate bases occur at one position. Comparison of any two genomes reveals around 1 SNP per kilobase. A sufficiently dense map of SNPs would allow the detection of sequence variants responsible for particular characteristics on the basis that they are associated with a specific SNP allele. Here we have evaluated large-scale sequencing approaches to obtaining SNPs, and have constructed a map of 2,730 SNPs on human chromosome 22. Most of the SNPs are within 25 kilobases of a transcribed exon, and are valuable for association studies. We have scaled up the process, detecting over 65,000 SNPs in the genome as part of The SNP Consortium programme, which is on target to build a map of 1 SNP every 5 kilobases that is integrated with the human genome sequence and that is freely available in the public domain.

Cell Line↗

Chromosome 20 deletions in myeloid malignancies: reduction of the common deleted region, generation of a PAC/BAC contig and identification of candidate genes. UK Cancer Cytogenetics Group (UKCCG).

Deletion of the long arm of chromosome 20 represents the most common chromosomal abnormality associated with the myeloproliferative disorders (MPDs) and is also found in other myeloid malignancies including myelodysplastic syndromes (MDS) and acute myeloid leukaemia (AML). Previous studies have identified a common deleted region (CDR) spanning approximately 8 Mb. We have now used G-banding, FISH or microsatellite PCR to analyse 113 patients with a 20q deletion associated with a myeloid malignancy. Our results define a new MPD CDR of 2.7 Mb, an MDS/AML CDR of 2.6 Mb and a combined 'myeloid' CDR of 1.7 Mb. We have also constructed the most detailed physical map of this region to date--a bacterial clone map spanning 5 Mb of the chromosome which contains 456 bacterial clones and 202 DNA markers. Fifty-one expressed sequences were localized within this contig of which 37 lie within the MPD CDR and 20 within the MDS/AML CDR. Of the 16 expressed sequences (six genes and 10 unique ESTs) within the 'myeloid' CDR, five were expressed in both normal bone marrow and purified CD34 positive cells. These data identify a set of genes which are both positional and expression candidates for the target gene(s) on 20q.

Antigens, CD34↗

Characterisation of a novel murine intestinal serine protease, DISP.

A putative novel murine serine protease, DISP, was identified by cDNA indexing and shown to be expressed primarily in distal gut. FISH analysis showed it to be localised to mouse chromosome 17A3. A possible human homologue for DISP has been identified. DISP is a novel member of clan SA/family S1 of the serine proteases, at present of unknown function.

Amino Acid Sequence↗

The Human Genome Project--an overview.

The human genome sequence will underpin human biology and medicine in the next century, providing a single, essential reference to all genetic information. The international program to determine the complete DNA sequence (3,000 million bases) is well underway. As of January 2000, 50% of the sequence is available in the public domain. A comprehensive working draft is expected this year, and the entire sequence is projected to be finished in 2003. DNA sequencing is carried out on mapped, overlapping bacterial clones of 150-200 kb. The working draft comprises assembled unfinished sequence and is released immediately in the public domain. The draft sequence of each clone is then completed, by closing any remaining gaps and resolving any ambiguities, before the entire sequence is checked, annotated, and submitted to the public databases. The sequence of each clone is finished to an accuracy of >99.99%. The availability of a reference sequence of the genome provides the basis for studying the nature of sequence variation, particularly single nucleotide polymorphisms (SNPs), in human populations. SNP typing is a powerful tool for genetic analysis, and will enable us to uncover the association of loci at specific sites in the genome with many disease traits. SNPs occur at a frequency of approximately 1 SNP/kb throughout the genome when the sequence of any two individuals is compared. Programs to detect and map SNPs in the human genome are underway with the aim of establishing a SNP map of the genome during the next two years. The human genome sequence will provide a complete description of all the genes. Annotation of the sequence with the gene structures is achieved by a combination of computational analysis (predictive and homology-based) and experimental confirmation by cDNA sequencing. Detecting homologies between newly defined gene products and proteins of known function helps to postulate biochemical functions for them, which can then be tested. Establishing the association of specific genes with disease phenotypes by mutation screening, particularly for monogenic disorders, provides further assistance in defining the functions of some gene products, as well as helping to establish the cause of the disease. As our knowledge of gene sequences and sequence variation in populations increases, we will pinpoint more and more of the genes and proteins that are important in common, complex diseases. A more detailed understanding of the function of the human genome will be achieved as we identify sequences that control gene expression. Given the availability of gene sequences, the expression status of genes in particular tissues can be monitored in parallel. By comparing corresponding genomic sequences in different species (for example: man, mouse, chicken, and zebrafish), regions that have been highly conserved during evolution can be identified, many of which reflect conserved functions such as gene regulation. These approaches promise to greatly accelerate our interpretation of the human genome sequence.

Human Genome Project↗

Analysis of vertebrate SCL loci identifies conserved enhancers.

The SCL gene encodes a highly conserved bHLH transcription factor with a pivotal role in hemopoiesis and vasculogenesis. We have sequenced and analyzed 320 kb of genomic DNA composing the SCL loci from human, mouse, and chicken. Long-range sequence comparisons demonstrated multiple peaks of human/mouse homology, a subset of which corresponded precisely with known SCL enhancers. Comparisons between mammalian and chicken sequences identified some, but not all, SCL enhancers. Moreover, one peak of human/mouse homology (+23 region), which did not correspond to a known enhancer, showed significant homology to an analogous region of the chicken SCL locus. A transgenic Xenopus reporter assay was established and demonstrated that the +23 region contained a new neural enhancer. This combination of long-range comparative sequence analysis with a high-throughput transgenic bioassay provides a powerful strategy for identifying and characterizing developmentally important enhancers.

Amino Acid Sequence↗

Decoding the human genome sequence.

The year 2000 is marked by the production of the sequence of the human genome. A 'working draft' of high quality sequence covering 90% of the genome has been determined and a quarter is in finished form, including the first two completed chromosomes. All sequence data from the project is made freely available to the community via the Internet, for further analysis and exploitation. The challenge which lies ahead is to decipher the information. Knowledge of the human genome sequence will enable us to understand how the genetic information determines the development, structure and function of the human body. We will be able to explore how variations within our DNA sequence cause disease, how they affect our interaction with our environment and ultimately to develop new and effective ways to improve human health.

Conserved Sequence↗

Improved method for detecting differentially expressed genes using cDNA indexing.

In cDNA indexing, differentially expressed genes are identified by the display of specific, corresponding subsets of cDNA. Subdivision of the cDNA population is achieved by the sequence-specific ligation of adapters to the overhangs created by class IIS restriction enzymes. However, inadequate specificity of ligation leads to redundancy between different adapter subsets. We evaluate the incidence of mismatches between adapters and class IIS restriction fragments during ligation and describe a modified set of conditions that improves ligation specificity. The improved protocol reduces redundancy between amplified cDNA subsets, which leads to a lower number of bands per lane of the differential display gel, and therefore simplifies analysis. We confirm the validity of this revised protocol by identifying five differentially expressed genes in mouse duodenum and ileum.

Animals↗

High-resolution landmark framework for the sequence-ready mapping of Xq23-q26.1.

We have established a landmark framework map over 20-25 Mb of the long arm of the human X chromosome using yeast artificial chromosome (YAC) clones. The map has approximately one landmark per 45 kb of DNA and stretches from DXS7531 in proximal Xq23 to DXS895 in proximal Xq26, connecting to published framework maps on its proximal and distal sides. There are three gaps in the framework map resulting from the failure to obtain clone coverage from the YAC resources available. Estimates of the maximum sizes of these gaps have been obtained. The four YAC contigs have been positioned and oriented using somatic-cell hybrids and fluorescence in situ hybridization, and the largest is estimated to cover approximately 15 Mb of DNA. The framework map is being used to assemble a sequence-ready map in large-insert bacterial clones, as part of an international effort to complete the sequence of the X chromosome. PAC and BAC contigs currently cover 18 Mb of the region, and from these, 12 Mb of finished sequence is available.

Blotting, Southern↗

A physical map of 30,000 human genes.

A map of 30,181 human gene-based markers was assembled and integrated with the current genetic map by radiation hybrid mapping. The new gene map contains nearly twice as many genes as the previous release, includes most genes that encode proteins of known function, and is twofold to threefold more accurate than the previous version. A redesigned, more informative and functional World Wide Web site (www.ncbi.nlm.nih.gov/genemap) provides the mapping information and associated data and annotations. This resource constitutes an important infrastructure and tool for the study of complex genetic traits, the positional cloning of disease genes, the cross-referencing of mammalian genomes, and validated human transcribed sequences for large-scale studies of gene expression.

Animals↗

A detailed physical and transcriptional map of the region of chromosome 20 that is deleted in myeloproliferative disorders and refinement of the common deleted region.

Acquired deletions of the long arm of chromosome 20 are the most common chromosomal abnormality seen in polycythemia vera and are also associated with other myeloid malignancies. Such deletions are believed to mark the site of one or more tumor suppressor genes, loss of which perturbs normal hematopoiesis. A common deleted region (CDR) has previously been identified on 20q. We have now constructed the most detailed physical map of this region to date--a YAC contig that encompasses the entire CDR and spans 23 cM (11 Mb). This contig contains 140 DNA markers and 65 unique expressed sequences. Our data represent a first step toward a complete transcriptional map of the CDR. The high marker density within the physical map permitted two complementary approaches to reducing the size of the CDR. Microsatellite PCR refined the centromeric boundary of the CDR to D20S465 and was used to search for homozygous deletions in 28 patients using 32 markers. No such deletions were detected. Genetic changes on the remaining chromosome 20 may therefore be too small to be detected or may occur in a subpopulation of cells.

Centromere↗

Host response to EBV infection in X-linked lymphoproliferative disease results from mutations in an SH2-domain encoding gene.

X-linked lymphoproliferative syndrome (XLP or Duncan disease) is characterized by extreme sensitivity to Epstein-Barr virus (EBV), resulting in a complex phenotype manifested by severe or fatal infectious mononucleosis, acquired hypogammaglobulinemia and malignant lymphoma. We have identified a gene, SH2D1A, that is mutated in XLP patients and encodes a novel protein composed of a single SH2 domain. SH2D1A is expressed in many tissues involved in the immune system. The identification of SH2D1A will allow the determination of its mechanism of action as a possible regulator of the EBV-induced immune response.

Antigens, CD↗

High-resolution physical map of the X-linked retinoschisis interval in Xp22.

X-linked retinoschisis (RS) is the leading cause of macular degeneration in young males and has been mapped to Xp22 between DXS418 and DXS999. To facilitate identification of the RS gene, we have constructed a yeast artificial chromosome (YAC) contig across this region comprising 28 YACs and 32 sequence-tagged sites including seven novel end clone markers. To establish the definitive marker order, a PAC contig containing 50 clones was also constructed, and all clones were fingerprinted. The marker order is: Xpter-DXS1317-(AFM205yd12-DXS7175-DXS7992) -60N8-T7-DXS1195-DXS7993-DXS7174 -60N8-SP6-DXS418-DXS7994-DXS7995-DXS7996-+ ++HYAT2-25HA10R-HYAT1-DXS7997-DXS7998- DXS257-434E8R-3542R-DXS6762-DXS7999-DXS 6763-434E8L-DXS8000-DXS6760-DXS7176- DXS8001-DXS999-3176R-PHKA2-Xcen. A long-range restriction map was constructed, and the RS region is estimated to be 1300 kb, containing three putative CpG islands. An unstable region was identified between DXS6763 and 434E8L. These data will facilitate positional cloning of RS and other disease genes in Xp22.

Chromosome Mapping↗