Search PubMed⌕ Search

Biomedical subjects

Ge Gao

Publications and source records attributed to Ge Gao.

23 records · Page 2Linked to original sources

PCAS--a precomputed proteome annotation database resource.

BACKGROUND: Many model proteomes or "complete" sets of proteins of given organisms are now publicly available. Much effort has been invested in computational annotation of those "draft" proteomes. Motif or domain based algorithms play a pivotal role in functional classification of proteins. Employing most available computational algorithms, mainly motif or domain recognition algorithms, we set up to develop an online proteome annotation system with integrated proteome annotation data to complement existing resources. RESULTS: We report here the development of PCAS (ProteinCentric Annotation System) as an online resource of pre-computed proteome annotation data. We applied most available motif or domain databases and their analysis methods, including hmmpfam search of HMMs in Pfam, SMART and TIGRFAM, RPS-PSIBLAST search of PSSMs in CDD, pfscan of PROSITE patterns and profiles, as well as PSI-BLAST search of SUPERFAMILY PSSMs. In addition, signal peptide and TM are predicted using SignalP and TMHMM respectively. We mapped SUPERFAMILY and COGs to InterPro, so the motif or domain databases are integrated through InterPro. PCAS displays table summaries of pre-computed data and a graphical presentation of motifs or domains relative to the protein. As of now, PCAS contains human IPI, mouse IPI, and rat IPI, A. thaliana, C. elegans, D. melanogaster, S. cerevisiae, and S. pombe proteome.PCAS is available at http://pak.cbi.pku.edu.cn/proteome/gca.php CONCLUSION: PCAS gives better annotation coverage for model proteomes by employing a wider collection of available algorithms. Besides presenting the most confident annotation data, PCAS also allows customized query so users can inspect statistically less significant boundary information as well. Therefore, besides providing general annotation information, PCAS could be used as a discovery platform. We plan to update PCAS twice a year. We will upgrade PCAS when new proteome annotation algorithms identified.

Algorithms↗

PepPat, a pattern-based oligopeptide homology search method and the identification of a novel tachykinin-like peptide.

UNLABELLED: PepPat, a hybrid method that combines pattern matching with similarity scoring, is described. We also report PepPat's application in the identification of a novel tachykinin-like peptide. PepPat takes as input a query peptide and a user-specified regular expression pattern within the peptide. It first performs a database pattern match and then ranks candidates on the basis of their similarity to the query peptide. PepPat calculates similarity over the pattern spanning region, enhancing PepPat's sensitivity for short query peptides. PepPat can also search for a user-specified number of occurrences of a repeated pattern within the target sequence. We illustrate PepPat's application in short peptide ligand mining. As a validation example, we report the identification of a novel tachykinin-like peptide, C14TKL-1, and show it is an NK1 (neuokinin receptor 1) agonist whose message is widely expressed in human periphery. AVAILABILITY: PepPat is offered online at: http://peppat.cbi.pku.edu.cn.

Algorithms↗

[Initial analysis of complete genome sequences of SARS coronavirus].

Multiple sequence alignment among 12 complete SARS coronavirus (SARS-CoV) sequences reveals that the major parts of 29708 b of the genomes have 99.82% identical bases. Forty two nucleotide mismatches were found in addition to the five and six gaps in two genomes. Among them, 28 mismatches result in changes of amino acid in the encoded proteins. Analysis of the changes implies possible effect on the Spike and Membrane protein of the virus, while most of the other changes seem not very significant to alter the structure and function of the proteins. These results have been released on the anti-sars web site maintained by the Centre of Bioinformatics, Peking University (antisars.cbi.pku.edu.cn) and may be of help for further experimental study.

Amino Acid Sequence↗

[Introduction to genome databases].

A brief introduction to the genome databases GDB, GenoList and Ensembl is given. These databases, mirrored and maintained at the Centre of Bioinformatics, Peking University, provide useful information for genome research.

English Abstract↗

Highly enantioselective phenylacetylene additions to both aliphatic and aromatic aldehydes.

The readily available and inexpensive BINOL in combination with Ti(O(i)Pr)(4) is found to catalyze the reaction of an alkynylzinc reagent with various types of aldehydes including aliphatic aldehydes, aromatic aldehydes, and other alpha,beta-unsaturated aldehydes to generate chiral propargyl alcohols with 91-99% ee at room temperature. No previous chiral catalysts have exhibited such a broad scope of enantioselectivity with respect to the type of aldehydes for this reaction. [reaction: see text]

Acetylene↗