Search PubMed⌕ Search

Biomedical subjects

Andrew E Firth

Publications and source records attributed to Andrew E Firth.

4 recordsLinked to original sources

Programmed ribosomal frameshifting during PLEKHM2 mRNA decoding generates a constitutively active proteoform that supports myocardial function.

Programmed ribosomal frameshifting is a process where a proportion of ribosomes change their reading frame on an mRNA. While frameshifting is commonly used by viruses, very few phylogenetically conserved examples are known in nuclear encoded genes. Here, we report a +1 frameshifting event during decoding of the human gene PLEKHM2 that provides access to a second internally overlapping ORF. The new carboxyl-terminal domain of this frameshift protein forms an α helix, which relieves PLEKHM2 from autoinhibition and allows it to move to the tips of cells without activation by ARL8. Reintroducing both the canonically translated and frameshifted protein are necessary to restore normal contractile function of PLEKHM2 knockout cardiomyocytes, demonstrating the necessity of frameshifting for normal cardiac activity.

Frameshifting, Ribosomal↗

Detecting overlapping coding sequences with pairwise alignments.

MOTIVATION: Overlapping gene coding sequences (CDSs) are particularly common in viruses but also occur in more complex genomes. Detecting such genes with conventional gene-finding algorithms can be difficult for several reasons. If an overlapping CDS is on the same read-strand as a known CDS, then there may not be a distinct promoter or mRNA. Furthermore, the constraints imposed by double-coding can result in atypical codon biases. However, these same constraints lead to particular mutation patterns that may be detectable in sequence alignments. RESULTS: In this paper, we investigate several statistics for detecting double-coding sequences with pairwise alignments--including a new maximum-likelihood method. We also develop a model for double-coding sequence evolution. Using simulated sequences generated with the model, we characterize the distribution of each statistic as a function of sequence composition, length, divergence time and double-coding frame. Using these results, we develop several algorithms for detecting overlapping CDSs. The algorithms were tested on known overlapping CDSs and other overlapping open reading frames (ORFs) in the hepatitis B virus (HBV), Escherichia coli and Salmonella typhimurium genomes. The algorithms should prove useful for detecting novel overlapping genes--especially short coding ORFs in viruses. AVAILABILITY: Programs may be obtained from the authors. SUPPLEMENTARY INFORMATION: http://biochem.otago.ac.nz/double.html.

Algorithms↗

User-friendly algorithms for estimating completeness and diversity in randomized protein-encoding libraries.

Directed evolution of proteins depends on the production of molecular diversity by random mutagenesis. While a number of methods have been developed for introducing this diversity, the best ways to sample it are not always clear. Here we used simple statistics to analyse completeness and diversity in randomized libraries generated by oligonucleotide-directed mutagenesis, error-prone polymerase chain reaction (epPCR) and in vitro recombination of highly homologous sequences. For oligonucleotide-directed mutagenesis, we derive equations to estimate how complete a given library is expected to be and also to predict the size of library required to give a fixed probability of being 100% complete. We describe the statistical bases for computer programs which estimate the number of distinct variants represented in epPCR and shuffled libraries, dubbed PEDEL and DRIVeR, respectively. These programs allow the user to calculate (rather than guess) the diversity represented in a given library and also provide empirical guidelines for maximizing this diversity. PEDEL and DRIVeR are available at www.bio.cam.ac.uk/ approximately blackburn/stats.html.

Algorithms↗