Search PubMed⌕ Search

PubMed · 10592270

Human immunodeficiency virus reverse transcriptase and protease sequence database.

Abstract

The HIV RT and Protease Sequence Database is an online relational database that catalogs evolutionary and drug-related human immunodeficiency virus (HIV) reverse transcriptase (RT) and protease sequence variation (http://hivdb.stanford.edu). The database contains a compilation of nearly all published HIV RT and protease sequences including International Collaboration database submissions (e.g., GenBank) and sequences published in journal articles. Sequences are linked to data about the source of the sequence sample and the antiretroviral drug treatment history of the individual from whom the isolate was obtained. The database is curated and sequences are annotated with data from >230 literature references. Users can retrieve additional data and view alignments of sequence sets meeting specific criteria (e.g., treatment history, subtype, presence of a particular mutation). A gene-specific sequence analysis program, new user-defined queries and nearly 2000 additional sequences were added in 1999.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

R W Shafer, D R Jung, B J Betts, Y Xi, M J Gonzales. 2000-01-01. Human immunodeficiency virus reverse transcriptase and protease sequence database.. https://doi.org/10.1093/nar%2F28.1.346

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Identification of protein domains on topological basis.

A theoretical method is proposed to identify structural domains in proteins of known structures. It is based on the distribution of the local axes of the polypeptide chain. In particular, a statistical analysis is applied to the contributions of the local axes to the absolute writhing number, a topological property of a space curve resulting from the number of self-crossings in the curve projections onto a unit sphere. This finding supports the hypothesis that topological requirements should be satisfied in the process of protein folding and in the final organization of the tertiary structures.

Databases, Factual↗

Using patient-reportable clinical history factors to predict myocardial infarction.

Using a derivation data set of 1253 patients, we built several logistic regression and neural network models to estimate the likelihood of myocardial infarction based upon patient-reportable clinical history factors only. The best performing logistic regression model and neural network model had C-indices of 0.8444 and 0.8503, respectively, when validated on an independent data set of 500 patients. We conclude that both logistic regression and neural network models can be built that successfully predict the probability of myocardial infarction based on patient-reportable history factors alone. These models could have important utility in applications outside of a hospital setting when objective diagnostic test information is not yet be available.

Databases, Factual↗

Why are "natively unfolded" proteins unstructured under physiologic conditions?

"Natively unfolded" proteins occupy a unique niche within the protein kingdom in that they lack ordered structure under conditions of neutral pH in vitro. Analysis of amino acid sequences, based on the normalized net charge and mean hydrophobicity, has been applied to two sets of proteins: small globular folded proteins and "natively unfolded" ones. The results show that "natively unfolded" proteins are specifically localized within a unique region of charge-hydrophobicity phase space and indicate that a combination of low overall hydrophobicity and large net charge represent a unique structural feature of "natively unfolded" proteins.

Databases, Factual↗