Search PubMed⌕ Search

SEARCH · Search PubMed

Results for “Data Accuracy”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 325 records · Page 18Linked to original sources

Molecular sequence accuracy: analysing imperfect data.

Molecular sequences are experimentally derived data that can be expected to contain errors as a result of diverse phenomena such as biological variation, molecular cloning artifacts, imperfect sequence determination, and data handling during contig assembly. Errors will affect the reliability of database searches and sequence alignments, but their impact may be minimized by the use of analytical techniques that anticipate that the data will be imperfect.

Amino Acid Sequence↗

Selecting diagnostic tests to identify febrile infants less than 3 months of age as being at low risk for serious bacterial infection: a scientific overview.

PURPOSE: To select diagnostic tests that confidently identify febrile infants less than 3 months of age seen at an outpatient facility as being at low risk for serious bacterial infection (SBI). DATA IDENTIFICATION: An English-language literature search employing MEDLINE (1966 to 1991), Science Citation Index (1977 to 1991) using key citations, bibliographic reviews of primary research and review articles, and correspondence with authors of recent articles. STUDY SELECTION: After independent review by two observers, 10 of 333 originally identified titles were selected on the basis of prespecified selection criteria. DATA EXTRACTION: Two observers independently assessed studies by using explicit methodologic criteria for evaluating the quality of studies dealing with diagnostic tests. One reviewer extracted all the data from the articles; the second reviewer checked these data for accuracy. RESULTS OF DATA ANALYSIS: On the basis of prespecified criteria, results were pooled from two studies that used the Rochester criteria, had high methodologic validity, and did not have significant heterogeneity (p = 0.32, Breslow-Day test), to give an estimate of the best negative likelihood ratio (95% confidence interval) for SBI = 0.03; 0 to 0.23). CONCLUSION: The negative likelihood ratio of 0.03 allowed us to conclude that after the Rochester criteria for low risk of SBI have been satisfied, the probability of SBI in a febrile infant less than 3 months of age drops from a baseline rate of 7% (or 1 in 14 infants) to 0.2% (or 1 in 500). An expectant approach in these low-risk infants is therefore a reasonable choice.

Bacterial Infections↗

Assessing Metal Ion Assignment Accuracy in Protein Data Bank Models via Elemental Spectroscopy.

Accurate representation of metal ions in macromolecular structures is critical for chemical interpretation, computational modeling, and machine-learning methods that rely on Protein Data Bank (PDB) entries. However, the elemental identity of metals modeled in crystallographic structures is often inferred indirectly and rarely validated experimentally. Here, we combine Particle Induced X-ray Emission (PIXE) and X-ray Fluorescence Spectroscopy (XRFS) to determine the elemental composition of protein samples used to generate 70 deposited metalloprotein crystal structures. By analyzing the original protein material employed for crystallization, but before the addition of crystallization buffer solutions, we assess whether the modeled metal ions in deposited structures are consistent with experimentally detectable elemental content. We find that in a majority of cases, the metals modeled in the corresponding PDB entries are inconsistent with the metals present in the protein samples before crystallization, or that additional metals are present but not represented in the structural models. Spectroscopic results were integrated with automated crystallographic validation metrics, including real-space Z-difference (RSZD) analysis and systematic rerefinement, to evaluate atomic-number mismatch at metal sites. PIXE and XRFS show strong agreement for dominant elemental signals and provide complementary, scalable approaches for identifying suspect metal assignments. This work does not address physiological or functional metalation but instead highlights a widespread data integrity issue in deposited macromolecular structures, PDB-wide. These results establish an experimentally corroborated link between elemental identity and crystallographic validation metrics, enabling the large-scale detection of chemically inconsistent annotations in structural databases used for computational modeling and machine learning.

Databases, Protein↗

Completeness and accuracy of interview data from proxy respondents: demographic, medical, and life-style factors.

To evaluate the quality of exposure data provided by proxy respondents, we used a dual interview protocol in a case-control study of subarachnoid hemorrhage. All control subjects and their proxy respondents were interviewed (N = 283 control-proxy pairs), as were the cases who were able to provide their own information and their proxy respondents (N = 68 case-proxy pairs). The reliability of proxy-derived data was excellent for demographic and body habitus measures (kappa or intraclass correlation range = 0.86-0.99), and all aspects of cigarette smoking history (range = 0.79-0.93). Proxy reliability was somewhat lower for questions regarding medications and hormone preparations (range = 0.55-0.88), alcohol consumption (range = 0.52-0.82), and recreational physical activity (range = 0.55-0.67). Proxy reliability varied according to the relationship of the proxy to the index subject. Relative to the index subjects, proxy respondents tended to underreport the presence or level of exposure. For most exposures, odds ratios computed with proxy-derived data were similar in magnitude to odds ratios obtained with index subject data; important bias due to differential nonresponse or differential misclassification was suggested only for questions regarding hormone replacement therapy. Epidemiologic studies that rely on proxy respondents may require more subjects to offset the effect of nondifferential nonresponse and misclassification on the precision of effect estimates.

Adolescent↗

Capturing and using clinical outcome data: implications for information systems design.

There is an urgent need to capture and record data related to clinical outcomes, but there are many barriers. The range of problems includes lack of agreement on conceptualization of the term "outcome," inadequate measures of outcomes, and inadequate information systems to capture and manipulate data that would reflect outcomes. This article focuses on information system requirements to capture, store, and utilize clinical outcome data. For greatest accuracy, outcome data should be captured as close to the source as possible, including direct data capture from patients themselves and from their families. To make maximum use of outcome data, systems must be designed to 1) store data in multipurpose databases; 2) share data across different platforms; 3) link outcome data to other data that might influence or explain outcomes; 4) allow querying of the data by authorized personnel; and 5) protect patient confidentiality.

Decision Support Systems, Management↗

Accuracy of extrapolated data as a function of prior knowledge and regularization.

The prior discrete Fourier transform (PDFT) is a linear spectral estimator that provides a solution that is both data consistent and of minimum weighted norm through the use of a suitably designed Hilbert space. The PDFT has been successfully used in imaging applications to improve resolution and overcome the nonuniqueness associated with having only finitely many spectral measurements. With the use of an appropriate prior function, the resolution of the reconstructed image can be improved dramatically. We explore the ways in which some significant parameters affect the PDFT estimate. A relationship between estimated spectral values, prior knowledge, and regularization was examined. It allows one to assess the reliability of the estimated spectral values for a given choice of prior estimate and provides a means for optimizing PDFT-based estimators.

Journal Article↗

The accuracy of industry data from death certificates for workplace homicide victims.

This study compared death certificate data on usual industry for workplace homicide victims in five urban Texas counties, with medical examiners' data on the industries where victims were working when injured. The overall positive predictive value of the death certificate data was 72 per cent. Death certificate data on usual industry underestimated the number of victims working in high-risk industries when injured, partly because of victims whose usual industry was recorded as student, housewife, or military personnel.

Death Certificates↗