Search PubMed⌕ Search

PubMed · 9750196

Automated sequence preprocessing in a large-scale sequencing environment.

Abstract

A software system for transforming fragments from four-color fluorescence-based gel electrophoresis experiments into assembled sequence is described. It has been developed for large-scale processing of all trace data, including shotgun and finishing reads, regardless of clone origin. Design considerations are discussed in detail, as are programming implementation and graphic tools. The importance of input validation, record tracking, and use of base quality values is emphasized. Several quality analysis metrics are proposed and applied to sample results from recently sequenced clones. Such quantities prove to be a valuable aid in evaluating modifications of sequencing protocol. The system is in full production use at both the Genome Sequencing Center and the Sanger Centre, for which combined weekly production is approximately 100, 000 sequencing reads per week.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

M C Wendl, S Dear, D Hodgson, L Hillier. 1998. Automated sequence preprocessing in a large-scale sequencing environment.. https://doi.org/10.1101/gr.8.9.975

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Application of semi-automated metabolite identification software in the drug discovery process for rapid identification of metabolites and the cytochrome P450 enzymes responsible for their formation.

Rapid identification of metabolites of compound X using data dependent scan function of a quadrupole ion trap mass spectrometer and semi-automated metabolite identification software is described. Compound X is metabolized via monooxygenation and desmethylation. LC-ESI-MS spectra obtained, following incubations of Compound X with microsomes in the presence and absence of chemical inhibitors specific for CYP1A2, CYP3A4, CYP2D6, CYP2C9 and CYP2E1, were processed using semi-automated metabolite identification software to extract information and to identify the cytochrome P450 enzymes responsible for metabolite formation. Chemical inhibition data suggests that the primary cytochrome P450 (CYP450) isozyme responsible for the metabolism of compound X is CYP3A4 with a minor contribution from both CYP2D6 and CYP2E1. Additionally, neither CYP2C9 nor CYP1A2 appears to contribute to the metabolism of compound X.

Automation↗

Protein crystallization for genomics: towards high-throughput optimization techniques.

Protein crystallization has gained a new strategic and commercial relevance in the next phase of the genome projects, in which X-ray crystallography will play a major role. Considerable advances have been made in the automation of protein preparation and also in the X-ray analysis and bioinformatics stages once diffraction-quality crystals are available. These advances have not yet been matched by equally good methods for the crystallization process itself. In the area of crystallization, the main effort and resources are currently being invested into the automation of screening procedures to identify potential crystallization conditions. However, in spite of the ability to generate numerous trials, so far only a small percentage of the proteins produced have led to structure determinations. This is because screening in itself is not usually enough; it has to be complemented by an equally important procedure in crystal production, namely crystal optimization. In the rush towards structural genomics, optimization techniques have been somewhat neglected, mainly because it was hoped that large-scale screening alone would produce the desired results. In addition, optimization has relied on particular individual methods that are often difficult to automate and to adapt to high throughput. This article addresses a major gap in the field of structural genomics by describing practical ways of automating individual optimization methods in order to adapt them to high-throughput techniques.

Automation↗

Semi-automated liquid--liquid back-extraction in a 96-well format to decrease sample preparation time for the determination of dextromethorphan and dextrorphan in human plasma.

A semi-automated, 96-well based liquid-liquid back-extraction (LLE) procedure was developed and used for sample preparation of dextromethorphan (DEX), an active ingredient in many over-the-counter cough formulations, and dextrorphan (DOR), an active metabolite of DEX, in human plasma. The plasma extracts were analyzed by liquid chromatography-tandem mass spectrometry (LC-MS-MS). The analytes were isolated from human plasma using an initial ether extraction, followed by a back extraction from the ether into a small volume of acidified water. The acidified water isolated from the back extraction was analyzed directly by LC-MS-MS, eliminating the need for a dry down step. A liquid handling system was utilized for all aspects of liquid transfers during the LLE procedure including the transfer of samples from individual tubes into a 96-well format, preparation of standards, addition of internal standard and the addition and transfer of the extraction solvents. The semi-automated, 96-well based LLE procedure reduced sample preparation time by a factor of four versus a comparable manually performed LLE procedure.

Automation↗