Search PubMed⌕ Search

Biomedical subjects

R D Appel

Publications and source records attributed to R D Appel.

At least 19 recordsLinked to original sources

The 1999 SWISS-2DPAGE database update.

SWISS-2DPAGE (http://www.expasy.ch/ch2d/ ) is an annotated two-dimensional polyacrylamide gel electro-phoresis (2-DE) database established in 1993. The current release contains 24 reference maps from human and mouse biological samples, as well as from Saccharomyces cerevisiae, Escherichia coli and Dictyostelium discoideum origin. These reference maps have now 2824 identified spots, corresponding to 614 separate protein entries in the database, in addition to virtual entries for each SWISS-PROT sequence or any user-entered amino acids sequence. Last year improvements in the SWISS-2DPAGE database are as follows: three new maps have been created and several others have been updated; cross-references to newly built federated 2-DE databases have been added; new functions to access the data have been provided through the ExPASy proteomics server.

Animals↗

A molecular scanner to automate proteomic research and to display proteome images.

Identification and characterization of all proteins expressed by a genome in biological samples represent major challenges in proteomics. Today's commonly used high-throughput approaches combine two-dimensional electrophoresis (2-DE) with peptide mass fingerprinting (PMF) analysis. Although automation is often possible, a number of limitations still adversely affect the rate of protein identification and annotation in 2-DE databases: the sequential excision process of pieces of gel containing protein; the enzymatic digestion step; the interpretation of mass spectra (reliability of identifications); and the manual updating of 2-DE databases. We present a highly automated method that generates a fully annoated 2-DE map. Using a parallel process, all proteins of a 2-DE are first simultaneously digested proteolytically and electro-transferred onto a poly(vinylidene difluoride) membrane. The membrane is then directly scanned by MALDI-TOF MS. After automated protein identification from the obtained peptide mass fingerprints using PeptIdent software (http://www.expasy.ch/tools/peptident.html + ++), a fully annotated 2-D map is created on-line. It is a multidimensional representation of a proteome that contains interpreted PMF data in addition to protein identification results. This "MS-imaging" method represents a major step toward the development of a clinical molecular scanner.

Automation↗

The SWISS-2DPAGE database: what has changed during the last year.

SWISS-2DPAGE (http://www.expasy.ch/ch2d/) is an annotated two-dimensional polyacrylamide gel electrophoresis (2-D PAGE) database established in 1993. The current release contains 21 reference maps from human and mouse biological samples, as well as from Saccharomyces cerevisiae, Escherichia coli and Dictyostelium discoideum origin. These reference maps now have 2480 identified spots, corresponding to 528 separate protein entries in the database, in addition to virtual entries for each SWISS-PROT sequence. During the last year, the SWISS-2DPAGE has undergone major changes. Six new maps have been added, and new functions to access the data have been provided through the ExPASy server. Finally, an important change concerns the database funding source.

Animals↗

Modeling peptide mass fingerprinting data using the atomic composition of peptides.

The peptide mass fingerprinting technique is commonly used for identifying proteins analyzed by mass spectrometry (MS) after enzymatic digestion. Our goal is to build a theoretical model that predicts the mass spectra of such digestion products in order to improve the identification and characterization of proteins using this technique. We present here the first step towards a full MS model. We have modeled MS spectra using the atomic composition of peptides and evaluated the influence that this composition may have on the MS signals. Peptides deduced from the SWISS-PROT protein sequence database were used for the calculation. To validate the model, the variability of the peptide mass distribution in SWISS-PROT was compared to two theoretical, randomly generated databases. Functions have been built that describe the behavior of the isotopic distribution according to the mass of peptides. The variability of these functions was analyzed. In particular, the influence of sulfur was studied. This work, while representing only a first step in the construction of an MS model, yields immediate practical results, as the new isotopic distribution model significantly improves peak detection in MS spectra used by protein identification algorithms.

Databases, Factual↗

Improving protein identification from peptide mass fingerprinting through a parameterized multi-level scoring algorithm and an optimized peak detection.

We have developed a new algorithm to identify proteins by means of peptide mass fingerprinting. Starting from the matrix-assisted laser desorption/ionization-time-of-flight (MALDI-TOF) spectra and environmental data such as species, isoelectric point and molecular weight, as well as chemical modifications or number of missed cleavages of a protein, the program performs a fully automated identification of the protein. The first step is a peak detection algorithm, which allows precise and fast determination of peptide masses, even if the peaks are of low intensity or they overlap. In the second step the masses and environmental data are used by the identification algorithm to search in protein sequence databases (SWISS-PROT and/or TrEMBL) for protein entries that match the input data. Consequently, a list of candidate proteins is selected from the database, and a score calculation provides a ranking according to the quality of the match. To define the most discriminating scoring calculation we analyzed the respective role of each parameter in two directions. The first one is based on filtering and exploratory effects, while the second direction focuses on the levels where the parameters intervene in the identification process. Thus, according to our analysis, all input parameters contribute to the score, however with different weights. Since it is difficult to estimate the weights in advance, they have been computed with a generic algorithm, using a training set of 91 protein spectra with their environmental data. We tested the resulting scoring calculation on a test set of ten proteins and compared the identification results with those of other peptide mass fingerprinting programs.

Algorithms↗

Two-dimensional electrophoresis resources available from ExPASy.

This paper describes the set of two-dimensional electrophoresis (2-DE) resources currently available from the ExPASy proteomics Web server. These resources include the SWISS-2DPAGE database, 2-DE software packages, 2-DE technical and educational services, as well as indexes and search engines for 2-DE related sites over the Internet.

Databases, Factual↗

[Internet for physicians: a tool for today and tomorrow ].

The Internet is becoming more and more part of our habits regarding documentation and communication, thus limiting the frontiers to only that of the language. As in many other fields, the Internet is also present in the medical domain. The Internet is accessible to all, providing a technology which is simple and ergonomic and furthermore less costly. A growing number of individuals are indeed offering information, and the multiplication and diversification of documentary thus resulting renders the quality often questionable and the search for information difficult. This article firstly presents the Internet services with examples in the medical domain and more particularly in paediatrics. It then identifies the problems related to the Internet and further details three tools useful to find medical information and to surf on the Internet: Medline, a bibliographical reference search tool (from the National Library of Medicine-NLM); Medhunt, a search tool of the Health on the Net Foundation specialised in the health domain; and the HONcode, a code of ethics developed to homogenize medical information on the Internet.

Computer User Training↗

Protein identification with N and C-terminal sequence tags in proteome projects.

Genome sequences are available for increasing numbers of organisms. The proteomes (protein complement expressed by the genome) of many such organisms are being studied with two-dimensional (2D) gel electrophoresis. Here we have investigated the application of short N-terminal and C-terminal sequence tags to the identification of proteins separated on 2D gels. The theoretical N and C termini of 15, 519 proteins, representing all SWISS-PROT entries for the organisms Mycoplasma genitalium, Bacillus subtilis, Escherichia coli, Saccharomyces cerevisiae and human, were analysed. Sequence tags were found to be surprisingly specific, with N-terminal tags of four amino acid residues found to be unique for between 43% and 83% of proteins, and C-terminal tags of four amino acid residues unique for between 74% and 97% of proteins, depending on the species studied. Sequence tags of five amino acid residues were found to be even more specific. To utilise this specificity of sequence tags for protein identification, we created a world-wide web-accessible protein identification program, TagIdent (http://www.expasy.ch/www/tools.html), which matches sequence tags of up to six amino acid residues as well as estimated protein pI and mass against proteins in the SWISS-PROT database. We demonstrate the utility of this identification approach with sequence tags generated from 91 different E. coli proteins purified by 2D gel electrophoresis. Fifty-one proteins were unambiguously identified by virtue of their sequence tags and estimated pI and mass, and a further 11 proteins identified when sequence tags were combined with protein amino acid composition data. We conlcude that the TagIdent identification approach is best suited to the identification of proteins from prokaryotes whose complete genome sequences are available. The approach is less well suited to proteins from eukaryotes, as many eukaryotic proteins are not amenable to sequencing via Edman degradation, and tag protein identification cannot be unambiguous unless an organism's complete sequence is available.

Amino Acid Sequence↗

Current status of the SWISS-2DPAGE database.

The SWISS-2DPAGE database (http: //www.expasy.ch/ch2d/ch2d-top.html ) consists of two-dimensional polyacrylamide gel electrophoresis images, as well as textual descriptions of the proteins that have been identified on them. The current release contains 15 reference maps from human biological samples, as well as from Saccharomyces cerevisiae , Escherichia coli and Dictyostelium discoideum origin. These reference maps have 2088 identified spots, corresponding to 410 separate protein entries in the database, in addition to virtual entries for each SWISS-PROT sequence.

Animals↗

'98 Escherichia coli SWISS-2DPAGE database update.

The combination of two-dimensional polyacrylamide gel electrophoresis (2-D PAGE), computer image analysis and several protein identification techniques allowed the Escherichia coli SWISS-2DPAGE database to be established. This is part of the ExPASy molecular biology server accessible through the WWW at the URL address http://www.expasy.ch/ch2d/ch2d-top.html . Here we report recent progress in the development of the E. coli SWISS-2DPAGE database. Proteins were separated with immobilized pH gradients in the first dimension and sodium dodecyl sulfate-polyacrylamide gel electrophoresis in the second dimension. To increase the resolution of the separation and thus the number of identified proteins, a variety of wide and narrow range immobilized pH gradients were used in the first dimension. Micropreparative gels were electroblotted onto polyvinylidene difluoride membranes and spots were visualized by amido black staining. Protein identification techniques such as amino acid composition analysis, gel comparison and microsequencing were used, as well as a recently described Edman "sequence tag" approach. Some of the above identification techniques were coupled with database searching tools. Currently 231 polypeptides are identified on the E. coli SWISS-2DPAGE map: 64 have been identified by N-terminal microsequencing, 39 by amino acid composition, and 82 by sequence tag. Of 153 proteins putatively identified by gel comparison, 65 have been confirmed. Many proteins have been identified using more than one technique. Faster progress in the E. coli proteome project will now be possible with advances in biochemical methodology and with the completion of the entire E. coli genome.

Bacterial Proteins↗

Multiple parameter cross-species protein identification using MultiIdent--a world-wide web accessible tool.

Recent increases in the number of genome sequencing projects means that the amount of protein sequence in databases is increasing at an astonishing pace. In proteome studies, this is facilitating the identification of proteins from molecularly well-defined organisms. However, in studies of proteins from the majority of organisms, proteins must be identified by comparing analytical data to sequences in databases from other species. This process is known as cross-species protein identification. Here we present a new program, MultiIdent, which uses multiple protein parameters such as amino acid composition, peptide masses, sequence tags, estimated protein pI and mass, to achieve cross-species protein identification. The program is structured so that protein amino acid composition, which is highly conserved across species boundaries, first generates a set of candidate proteins. These proteins are then queried with other protein parameters such as sequence tags and peptide masses. A final list of database entries which considers all analytical parameters is presented, ranked by an integrated score. We illustrate the power of the approach with the identification of a set of standard proteins, and the identification of proteins from dog heart separated by two-dimensional gel electrophoresis. The MultiIdent program is available on the world-wide web at: http://www.expasy.ch/sprot/multiident.h tml.

Amino Acid Sequence↗

Trends in medical information retrieval on Internet.

Information on the World Wide Web is unstructured, distributed, multimedia and multilingual. Many tools have been developed to help users search for useful information: subject hierarchies, general search engines, browsers and search assistants. Although helpful, they present serious limitations, mainly in terms of precision, multilingual indexing and distribution. In this paper, we cover some on-line solutions to medical information discovery and present our own approach, the MARVIN (multi-agent retrieval vagabond on information network) project, which tackles medical information research with specialized cooperative retrieval agents. We also draw some outlines for future extensions.

Computer Systems↗

The Health On the Net Code of Conduct for medical and health Websites.

Internet has become one of the most used communication media. This and the fact that no constraining information publishing policy exists have created an urgent need to control the quality of information circulating through this media. To this purpose, the Health On the Net Foundation has initiated the Code of Conduct (HONcode) for the health/medical domain. This initiative proposes guidelines to information providers, with the aim, on the one hand, of raising the quality of data available on the Net and, on the other hand, of helping to identify Internet sites that are maintained by qualified people and contain reliable data. The HONcode mainly includes the following ethical aspects: the author's credentials, the date of the last modification with respect to clinical documents, confidentiality of data, source data reference, funding and the advertising policy. This article presents the HONcode and its evolution since it was launched in 1996.

Advertising↗