Search PubMed⌕ Search

Biomedical subjects

Thomas Lengauer

Publications and source records attributed to Thomas Lengauer.

At least 37 records · Page 2Linked to original sources

Multiple-ligand-based virtual screening: methods and applications of the MTree approach.

We present a novel approach for ligand-based virtual screening by combining query molecules into a multiple feature tree model called MTree. All molecules are described by the established feature tree descriptor, which is derived from a topological molecular graph. A new pairwise alignment algorithm leads to a consistent topological molecular alignment based on chemically reasonable matching of corresponding functional groups. These multiple feature tree models find application in ligand-based virtual screening to identify new lead structures for chemical optimization. Retrospective virtual screening with MTree models generated for angiotensin-converting enzyme and the alpha1a receptor on a large candidate database yielded enrichment factors up to 71 for the first 1% of the screened database. MTree models outperformed database searches using single feature trees in terms of hit rates and quality and additionally identified alternative molecular scaffolds not included in any of the query molecules. Furthermore, relevant molecular features, which are known to be important for affinity to the target, are identified by this new methodology.

Adrenergic alpha-1 Receptor Antagonists↗

Clinical significance of in vitro replication-enhancing mutations of the hepatitis C virus (HCV) replicon in patients with chronic HCV infection.

BACKGROUND: Mutations in nonstructural (NS) hepatitis C virus (HCV) proteins enhance replication in HCV-1a/b replicons. The prevalence of such mutations and their clinical significance in vivo are unknown. METHODS: Parts of HCV NS3 and NS4B-NS5B genes that included 31 in vitro replication-enhancing sites were sequenced for 26 patients with chronic HCV genotype 1 infection. RESULTS: Five patients showed specific mutations within NS3 at sites enhancing replication in the replicon. Those mutations were associated with a slower decrease in HCV RNA concentration during interferon (IFN)- alpha -based therapy (P = .007). Neither specific nor other mutations within NS3 and NS4B-NS5B were associated with baseline HCV RNA concentrations. Within NS5A, fewer mutations in the major HCV strain (P = .001) and increased quasi-species complexity (P = .02) and diversity (P = .02) correlated with increasing baseline HCV RNA concentrations. In silico analyses of NS3 protein structures suggested that the majority of observed mutations did not lead to major conformational changes. CONCLUSIONS: Specific mutations leading to enhanced replication in the replicon system were detected in 5 of 26 patients in vivo and were not associated with baseline HCV RNA concentrations but were associated with a slower decrease in HCV RNA concentration during IFN- alpha -based therapy. Quasi-species heterogeneity of NS5A correlated with baseline HCV RNA concentrations.

Adult↗

Computational methods for the design of effective therapies against drug resistant HIV strains.

The development of drug resistance is a major obstacle to successful treatment of HIV infection. The extraordinary replication dynamics of HIV facilitates its escape from selective pressure exerted by the human immune system and by combination drug therapy. We have developed several computational methods whose combined use can support the design of optimal antiretroviral therapies based on viral genomic data.

Database Management Systems↗

Decomposing protein networks into domain-domain interactions.

UNLABELLED: The application of novel experimental techniques has generated large networks of protein-protein interactions. Frequently, important information on the structure and cellular function of protein-protein interactions can be gained from the domains of interacting proteins. We have designed a Cytoscape plugin that decomposes interacting proteins into their respective domains and computes a putative network of corresponding domain-domain interactions. To this end, the network graph of proteins has been extended by additional node and edge types for domain interactions, including different node and edge shapes and coloring schemes used for visualization. An additional plugin provides supplementary web links to Internet resources on domain function and structure. AVAILABILITY: Both Cytoscape plugins can be downloaded from http://www.cytoscape.org

Algorithms↗

BiQ Analyzer: visualization and quality control for DNA methylation data from bisulfite sequencing.

SUMMARY: Manual processing of DNA methylation data from bisulfite sequencing is a tedious and error-prone task. Here we present an interactive software tool that provides start-to-end support for this process. In an easy-to-use manner, the tool helps the user to import the sequence files from the sequencer, to align them, to exclude or correct critical sequences, to document the experiment, to perform basic statistics and to produce publication-quality diagrams. Emphasis is put on quality control: The program automatically assesses data quality and provides warnings and suggestions for dealing with critical sequences. The BiQ Analyzer program is implemented in the Java programming language and runs on any platform for which a recent Java virtual machine is available. AVAILABILITY: The program is available without charge for non-commercial users and can be downloaded from http://biq-analyzer.bioinf.mpi-inf.mpg.de/

DNA↗

Dissection of the inflammatory bowel disease transcriptome using genome-wide cDNA microarrays.

BACKGROUND: The differential pathophysiologic mechanisms that trigger and maintain the two forms of inflammatory bowel disease (IBD), Crohn disease (CD), and ulcerative colitis (UC) are only partially understood. cDNA microarrays can be used to decipher gene regulation events at a genome-wide level and to identify novel unknown genes that might be involved in perpetuating inflammatory disease progression. METHODS AND FINDINGS: High-density cDNA microarrays representing 33,792 UniGene clusters were prepared. Biopsies were taken from the sigmoid colon of normal controls (n = 11), CD patients (n = 10) and UC patients (n = 10). 33P-radiolabeled cDNA from purified poly(A)+ RNA extracted from biopsies (unpooled) was hybridized to the arrays. We identified 500 and 272 transcripts differentially regulated in CD and UC, respectively. Interesting hits were independently verified by real-time PCR in a second sample of 100 individuals, and immunohistochemistry was used for exemplary localization. The main findings point to novel molecules important in abnormal immune regulation and the highly disturbed cell biology of colonic epithelial cells in IBD pathogenesis, e.g., CYLD (cylindromatosis, turban tumor syndrome) and CDH11 (cadherin 11, type 2). By the nature of the array setup, many of the genes identified were to our knowledge previously uncharacterized, and prediction of the putative function of a subsection of these genes indicate that some could be involved in early events in disease pathophysiology. CONCLUSION: A comprehensive set of candidate genes not previously associated with IBD was revealed, which underlines the polygenic and complex nature of the disease. It points out substantial differences in pathophysiology between CD and UC. The multiple unknown genes identified may stimulate new research in the fields of barrier mechanisms and cell signalling in the context of IBD, and ultimately new therapeutic approaches.

Adolescent↗

Ataxin-2 and huntingtin interact with endophilin-A complexes to function in plastin-associated pathways.

Spinocerebellar ataxia type 2 is an inherited neurodegenerative disorder that is caused by an expanded trinucleotide repeat in the SCA2 gene, encoding a polyglutamine stretch in the gene product ataxin-2. Although evidence has been provided that ataxin-2 is involved in RNA metabolism, the physiological function of ataxin-2 remains unclear. Here, we demonstrate that ataxin-2 interacts with two members of the endophilin family, endophilin-A1 and endophilin-A3. To elucidate the physiological implications of these interactions, we exploited yeast as a model system and discovered that expression of ataxin-2 as well as both endophilin proteins is toxic for yeast lacking the SAC6 gene product fimbrin, a protein involved in actin filament organization and endocytotic processes. Intriguingly, expression of huntingtin, another polyglutamine protein interacting with endophilin-A3, was also toxic in Deltasac6 yeast. These effects can be suppressed by simultaneous expression of one of the two human fimbrin orthologs, L- or T-plastin. Moreover, we have discovered that ataxin-2 associates with L- and T-plastin and that overexpression of ataxin-2 leads to accumulation of T-plastin in mammalian cells. Thus, our findings suggest an interplay between ataxin-2, endophilin proteins and huntingtin in plastin-associated cellular pathways.

Adaptor Proteins, Signal Transducing↗

ROCR: visualizing classifier performance in R.

UNLABELLED: ROCR is a package for evaluating and visualizing the performance of scoring classifiers in the statistical language R. It features over 25 performance measures that can be freely combined to create two-dimensional performance curves. Standard methods for investigating trade-offs between specific performance measures are available within a uniform framework, including receiver operating characteristic (ROC) graphs, precision/recall plots, lift charts and cost curves. ROCR integrates tightly with R's powerful graphics capabilities, thus allowing for highly adjustable plots. Being equipped with only three commands and reasonable default values for optional parameters, ROCR combines flexibility with ease of usage. AVAILABILITY: http://rocr.bioinf.mpi-sb.mpg.de. ROCR can be used under the terms of the GNU General Public License. Running within R, it is platform-independent. CONTACT: tobias.sing@mpi-sb.mpg.de.

Computer Graphics↗

Structural and functional analysis of a novel mutation of CYP21B in a heterozygote carrier of 21-hydroxylase deficiency.

Congenital adrenal hyperplasia (CAH) due to 21-hydroxylase deficiency is one of the most common autosomal recessive disorders and occurs in its non-classical form in up to 6% of hirsute women. We report on a young woman with the clinical diagnosis of non-classical CAH and a novel, heterozygous missense mutation CTG-->GTG in exon 8, codon 317, of the steroid 21-hydroxylase CYP21B and complete loss of pseudogenes. Protein sequences of closely related P450 cytochromes and a homology-based 3D model of CYP21B were used for further functional analyses. We found that the mutated residue is part of a large cluster of hydrophobic residues. This cluster has three important features: (1) it is located directly next to the binding pocket, in close vicinity of the heme-cofactor, (2) all amino acids of the cluster are directly connected to two important binding regions, and (3) the packing within the cluster is very dense. Due to the tight packing in the cluster and its direct connection to the binding pocket region, any changes induced by the mutation of residue 317 can be expected to lead to structural shifts within the binding pocket and can explain the clinically observed impairment of 21-hydroxylase activity. In conclusion, the novel mutation L317V of the steroid 21-hydroxylase gene is associated with reduced steroid 21-hydroxylase activity probably due to structural shifts within the binding pocket and a mild phenotype of steroid 21-hydroxylase deficiency. In addition, the results support previous findings in which heterozygous CYP21 mutations are associated with symptoms of hyperandrogenism in susceptible individuals.

Adolescent↗

Confirmation of human protein interaction data by human expression data.

BACKGROUND: With microarray technology the expression of thousands of genes can be measured simultaneously. It is well known that the expression levels of genes of interacting proteins are correlated significantly more strongly in Saccharomyces cerevisiae than those of proteins that are not interacting. The objective of this work is to investigate whether this observation extends to the human genome. RESULTS: We investigated the quantitative relationship between expression levels of genes encoding interacting proteins and genes encoding random protein pairs. Therefore we studied 1369 interacting human protein pairs and human gene expression levels of 155 arrays. We were able to establish a statistically significantly higher correlation between the expression levels of genes whose proteins interact compared to random protein pairs. Additionally we were able to provide evidence that genes encoding proteins belonging to the same GO-class show correlated expression levels. CONCLUSION: This finding is concurrent with the naive hypothesis that the scales of production of interacting proteins are linked because an efficient interaction demands that involved proteins are available to some degree. The goal of further research in this field will be to understand the biological mechanisms behind this observation.

Cluster Analysis↗

Estimating HIV evolutionary pathways and the genetic barrier to drug resistance.

BACKGROUND: The evolution of drug-resistant viruses challenges the management of human immunodeficiency virus (HIV) infections. Understanding this evolutionary process is important for the design of effective therapeutic strategies. METHODS: We used mutagenetic trees, a family of probabilistic graphical models, to describe the accumulation of resistance-associated mutations in the viral genome. On the basis of these models, we defined the genetic barrier, a quantity that summarizes the difficulty for the virus to escape from the selective pressure of the drug by developing escape mutations. RESULTS: From HIV reverse-transcriptase sequences that had been obtained from treated patients, we derived evolutionary models for zidovudine, zidovudine plus lamivudine, and zidovudine plus didanosine. The genetic barriers to resistance to zidovudine, stavudine, lamivudine, and didanosine, for the above 3 regimens, were computed and analyzed. We found both the mode and the rate of development of resistance to be heterogeneous. The genetic barrier to zidovudine resistance was increased if lamivudine was added to zidovudine but was decreased for didanosine. The barrier to lamivudine resistance was maintained with zidovudine plus didanosine, whereas the barrier to didanosine resistance was reduced most with zidovudine plus lamivudine. CONCLUSION: Mutagenetic trees provide a quantitative picture of the evolution of drug resistance. The genetic barrier is a useful tool for design of effective treatment strategies.

Anti-HIV Agents↗

Synthesis and evaluation of imidazolylmethylenetetrahydronaphthalenes and imidazolylmethyleneindanes: potent inhibitors of aldosterone synthase.

Elevated plasma aldosterone levels play a detrimental role in certain forms of congestive heart failure and myocardial fibrosis. We proposed aldosterone synthase (CYP11B2) as a novel target for the treatment of these diseases. In this study, the synthesis and biological evaluation of substituted E- and Z-imidazolylmethylenetetrahydronaphthalenes and E- and Z-imidazolylmethyleneindanes (compounds 1a,b-9a,b) is described. The compounds were prepared by a Wittig-like reaction. They were tested for activity using bovine CYP11B and human CYP11B2 expressed in fission yeast and V79 MZh cells. Selectivity was determined toward human CYP11B1, CYP19, and CYP17. Especially in the case of CYP11B1 (steroid 11beta-hydroxylase), selectivity is a crucial issue, since sequence homology between this enzyme and the target enzyme is very high (93%). On the basis of the X-ray structure of human CYP2C9, a protein model of CYP11B2 was developed and docking experiments with the title compounds were performed. The biological results revealed highly potent inhibitors of CYP11B2 (IC(50) = 4-93 nM). The Z-isomers usually were more active than the corresponding E-isomers. Different inhibitory profiles could be observed: rather selective inhibitors of CYP11B1, dual inhibitors of both enzymes, and rather selective inhibitors of CYP11B2. The chloro derivative 8b was found to be a highly potent CYP11B2 inhibitor (IC(50) = 4 nM) showing a 5-fold selectivity for CYP11B1 (IC(50) = 20 nM). This compound could be an interesting lead for further optimization as a therapeutic agent. It also could be used as well as the CYP11B1 selective compounds as a pharmacological tool.

Animals↗

Synthesis and evaluation of (pyridylmethylene)tetrahydronaphthalenes/-indanes and structurally modified derivatives: potent and selective inhibitors of aldosterone synthase.

Elevated aldosterone levels are key effectors for the development and progression of congestive heart failure and myocardial fibrosis. Recently, we proposed inhibition of aldosterone synthase (CYP11B2) as an innovative strategy for the treatment of these diseases. In this study, the synthesis and biological evaluation of E- and Z-(pyridylmethylene)tetrahydronaphthalenes and -indanes (1a,b-38a) is described. The activity of the compounds was determined using human CYP11B2, and the selectivity was evaluated toward the human steroidogenic enzymes CYP11B1, CYP19, and CYP17. The biological results revealed a few rather selective inhibitors of CYP11B1, some compounds inhibiting both CYP11B1 and CYP11B2, and a large number of highly selective inhibitors of CYP11B2. The most active inhibitor was the 3-pyridyl compound 5a (IC(50) = 7 nM). The pyrimidyl-substituted derivative 28a was found to be the most selective CYP11B2 inhibitor (IC(50) = 27 nM) in this series, showing a 120-fold selectivity for CYP11B1 (IC(50) = 3179 nM). Molecular modeling, i.e., examination of the electronic and steric features of selected compounds and homology modeling and docking, was used to understand the structure-activity/-selectivity relationships.

Adrenal Cortex Hormones↗

Sarcoidosis is associated with a truncating splice site mutation in BTNL2.

Sarcoidosis is a polygenic immune disorder with predominant manifestation in the lung. Genome-wide linkage analysis previously indicated that the extended major histocompatibility locus on chromosome 6p was linked to susceptibility to sarcoidosis. Here, we carried out a systematic three-stage SNP scan of 16.4 Mb on chromosome 6p21 in as many as 947 independent cases of familial and sporadic sarcoidosis and found that a 15-kb segment of the gene butyrophilin-like 2 (BTNL2) was associated with the disease. The primary disease-associated variant (rs2076530; P(TDT) = 3 x 10(-6), P(case-control) = 1.1 x 10(-8); replication P(TDT) = 0.0018, P(case-control) = 1.8 x 10(-6)) represents a risk factor that is independent of variation in HLA-DRB1. BTNL2 is a member of the immunoglobulin superfamily and has been implicated as a costimulatory molecule involved in T-cell activation on the basis of its homology to B7-1. The G --> A transition constituting rs2076530 leads to the use of a cryptic splice site located 4 bp upstream of the affected wild-type donor site. Transcripts of the risk-associated allele have a premature stop in the spliced mRNA. The resulting protein lacks the C-terminal IgC domain and transmembrane helix, thereby disrupting the membrane localization of the protein, as shown in experiments using green fluorescent protein and V5 fusion proteins.

Bronchoalveolar Lavage↗

The HIN domain of IFI-200 proteins consists of two OB folds.

The interferon-inducible p200 (IFI-200/HIN-200) family of proteins regulates cell growth and differentiation, and confers resistance to the development of tumors and virus infections. IFI-200 family members are thought to exert their biological effects by modulation of the transcriptional activities of numerous factors and interaction with other proteins through the C-terminal HIN domains. However, the HIN domain structure and function have remained obscure. Therefore, we performed a comprehensive bioinformatics analysis and assembled a structure-based multiple sequence alignment of IFI-200 proteins. The application of fold recognition methods revealed that the HIN domain consists of two consecutive OB domains. Our structural models of DNA-binding HIN domains afford the long-sought interpretations for many previous experimental observations. Our results also raise the possibility of as yet unexplored functional roles of IFI-200 proteins as transcriptional regulators and as interaction partners of proteins involved in immunomodulatory and apoptotic processes.

Apoptosis↗

Estimating cancer survival and clinical outcome based on genetic tumor progression scores.

MOTIVATION: In cancer research, prediction of time to death or relapse is important for a meaningful tumor classification and selecting appropriate therapies. Survival prognosis is typically based on clinical and histological parameters. There is increasing interest in identifying genetic markers that better capture the status of a tumor in order to improve on existing predictions. The accumulation of genetic alterations during tumor progression can be used for the assessment of the genetic status of the tumor. For modeling dependences between the genetic events, evolutionary tree models have been applied. RESULTS: Mixture models of oncogenetic trees provide a probabilistic framework for the estimation of typical pathogenetic routes. From these models we derive a genetic progression score (GPS) that estimates the genetic status of a tumor. GPS is calculated for glioblastoma patients from loss of heterozygosity measurements and for prostate cancer patients from comparative genomic hybridization measurements. Cox proportional hazard models are then fitted to observed survival times of glioblastoma patients and to times until PSA relapse following radical prostatectomy of prostate cancer patients. It turns out that the genetically defined GPS is predictive even after adjustment for classical clinical markers and thus can be considered a medically relevant prognostic factor. AVAILABILITY: Mtreemix, a software package for estimating tree mixture models, is freely available for non-commercial users at http://mtreemix.bioinf.mpi-sb.mpg.de. The raw cancer datasets and R code for the analysis with Cox models are available upon request from the corresponding author.

Biomarkers, Tumor↗

Mtreemix: a software package for learning and using mixture models of mutagenetic trees.

SUMMARY: Mixture models of mutagenetic trees constitute a class of probabilistic models for describing evolutionary processes that are characterized by the accumulation of permanent genetic changes. They have been applied to model the accumulation of chromosomal gains and losses in tumor development and the development of drug resistance-associated mutations in the HIV genome.Mtreemix is a software package for estimating mutagenetic trees mixture models from observed cross-sectional data and for using these models for predictions. We provide programs for model fitting, model selection, simulation, likelihood computation and waiting time estimation. AVAILABILITY: Mtreemix, including source code, documentation, sample data files and precompiled Solaris and Linux binaries, is freely available for non-commercial users at http://mtreemix.bioinf.mpi-sb.mpg.de/

Algorithms↗

A new CARD15 mutation in Blau syndrome.

The caspase recruitment domain gene CARD15/NOD2, encoding a cellular receptor involved in an NF-kappaB-mediated pathway of innate immunity, was first identified as a major susceptibility gene for Crohn's disease (CD), and more recently, as responsible for Blau syndrome (BS), a rare autosomal-dominant trait characterized by arthritis, uveitis, skin rash and granulomatous inflammation. While CARD15 variants associated with CD are located within or near the C-terminal leucine-rich repeat domain and cause decreased NF-kappaB activation, BS mutations affect the central nucleotide-binding NACHT domain and result in increased NF-kappaB activation. In an Italian family with BS, we detected a novel mutation E383K, whose pathogenicity is strongly supported by cosegregation with the disease in the family and absence in controls, and by the evolutionary conservation and structural role of the affected glutamate close to the Walker B motif of the nucleotide-binding site in the NACHT domain. Interestingly, substitutions at corresponding positions in another NACHT family member cause similar autoinflammatory phenotypes.

Amino Acid Sequence↗