Search PubMedSearch

SEARCH · Search PubMed

Results for “data quality”

Search indexed PubMed citations on genomics, clinical trials, systematic reviews and public health. Explore titles, authors and supplied subject terms, then open the PubMed record.

Quote a phrase for an exact phrase match. Source license links do not imply unrestricted reuse.

At least 19 recordsLinked to original sources

An investigation into the distribution of radial immunodiffusion quality control data.

Quality control data from routine radial immunodiffusion assays for IgG, IgA, IgM, C3, C4 and alpha1-antitrypsin were tested by the Kolmogorov--Smirnov procedure for gaussian distribution. All but alpha1-antitrypsin were nongaussian in type. Further analysis of these date by plotting on log-normal probability paper showed them to have a log-normal distribution. Treatment of the date by either gaussian or nonparametric statistical methods produced little difference in confidence limits. It does not appear necessary to use nonparametric methods to calculate confidence limits from quality control data for the procedures studied.

Complement C3

Establishing a national pediatric stem cell transplantation registry in Iran addressing implementation and data quality challenges.

The Iranian Pediatric Hematopoietic Stem Cell Transplantation Registry (IPED-HSCT) was established to enhance data collection, improve patient outcomes, and support clinical research in pediatric hematopoietic stem cell transplantation. This study aimed to assess the feasibility and reliability of implementing a standardized registry in pediatric settings. A community-based participatory study was conducted across three pediatric HSCT centers in Iran. The registry development involved a multi-phase approach, including pilot testing and the implementation of a web-based system. Data were collected from fifty pediatric patients who underwent HSCT for both malignant and non-malignant conditions, with a focus on data completeness and user satisfaction. Statistical analyses were performed using IBM SPSS Statistics. The registry achieved a data completeness rate exceeding 90%, with a participant demographic of 31 males (62%) and 19 females (38%). Rigorous quality control measures and real-time validation rules were implemented, enhancing data reliability. User feedback indicated high satisfaction with the platform's design and training sessions. Challenges included variations in long-term follow-up data collection across centers. The IPED-HSCT Registry demonstrates that establishing a robust pediatric HSCT registry is feasible even in resource-limited settings. Its innovative features offer a scalable model for similar initiatives in developing countries. Future research should focus on ensuring long-term sustainability and fostering international collaborations to improve pediatric HSCT outcomes globally. not applicable.

Humans

Evaluating Wearable Devices for Remote Monitoring in Psychosis: Pilot Study Nested Within the CONNECT Cohort Study.

BACKGROUND: Digital remote monitoring technologies, including smartphones and wearables, offer promising avenues for early detection of psychosis relapse. However, selecting devices that are acceptable to participants and produce high-quality data remains challenging. OBJECTIVE: The aim of this nested pilot study was to assess the acceptability and data quality of 3 commercially available wearable devices in people with psychosis recruited to the CONNECT cohort study. METHODS: Participants recruited to the CONNECT study before July 31, 2024, were included in the pilot study and selected 1 of 3 wearable devices: a Fitbit Charge 5, Samsung Galaxy Watch 5, or Apple Watch SE. Baseline demographics were compared between device groups. Acceptability of devices to participants was assessed through a Wearable Device Satisfaction Questionnaire after 3 months of use, with the proportion of positive responses to each question calculated and compared. Data completeness was also assessed by calculating the number (and percentage) of valid days of step count, heart rate, and sleep data, and comparing between groups. Data quality was assessed through summarizing the amount of troubleshooting required, additional metrics available from the wearables, and continuity of data completeness by calculating the proportion of participants with at least 3 days of heart rate data per week for the first 20 weeks of follow-up. Predefined criteria were used to determine the next steps for the wider CONNECT study: if one device was superior, this would be selected; if none were found to be superior and the Fitbit was found to be noninferior, then Fitbit would be retained. RESULTS: Of the first 107 participants recruited to CONNECT, 105 were included in the pilot study evaluation. The Samsung Galaxy Watch was selected most frequently by participants (46/105, 43.8%), followed by the Apple Watch (27/105, 25.7%), and Fitbit Charge (23/105, 21.9%). Differences in participant demographics were observed across device groups. Self-reported acceptability after use did not differ substantially between devices. However, in terms of data completeness, the median proportion of valid heart rate data days was significantly lower for Samsung Galaxy (median 31.2%, IQR 8.5%-46.0%) compared to Fitbit (median 80.1%, IQR 26.7%-95.0%; P=.003) and Apple Watch (median 49.3%, IQR 21.5%-86.0%; P=.02). There was no significant difference between Fitbit and Apple Watch. Similar patterns were observed for step count and sleep data. The Samsung Galaxy Watch required more frequent troubleshooting for data flow issues and lacked additional physiological metrics, available from the other devices. CONCLUSIONS: Due to comparatively lower data quality and technical performance, the Samsung Galaxy Watch was discontinued for use in the subsequent phase of the CONNECT study. The study highlights the importance of incorporating nested evaluations of devices in long-term research.

Humans

ATAC-seq in Emerging Model Organisms: Challenges and Strategies.

The Assay for Transposase-Accessible Chromatin with sequencing (ATAC-seq) is a versatile and widely utilized method for identifying potential regulatory regions, such as promoters and enhancers, within a genome. ATAC-seq has been successfully applied to a wide range of established and emerging model organisms. However, implementing this method in emerging model systems, such as arthropods, can be challenging due to several factors that influence data quality. These factors include the availability of a sufficient amount and quality of tissue or cells, the need for species- and tissue-specific protocol optimization, the completeness and accuracy of the reference genome, and the quality of the genome annotation. In this article, we emphasize the key steps in the ATAC-seq protocol that, based on our experience, have the greatest impact on data quality when adapting this method for emerging model organisms. Specifically, we discuss the importance of nuclei isolation, the incubation conditions of the Tn5 transposase, and PCR amplification of the library. Furthermore, we outline essential quality checkpoints during the bioinformatic analysis of ATAC-seq data to assist in assessing data integrity and consistency. Given that many emerging model organisms may not be readily available in laboratory cultures, we also emphasize the importance of evaluating how different preservation methods affect ATAC-seq data quality. Based on examples in one spider and one ant species, we demonstrate that replication and thorough quality controls at all steps of the protocol and data analysis are essential to assess the usability of ATAC-seq data. Our data highlights the importance of isolating the right number of intact nuclei, as well as ensuring optimal amplification conditions during library preparation to obtain good-quality sequence data for downstream analyses. We recommend using fresh tissue samples if possible because we show that direct cryopreservation of the tissue may affect chromatin integrity. This effect could be avoided or reduced by preserving the homogenate in cell culture medium. Overall, we explain the ATAC-seq protocol and downstream analyses in detail and give step-by-step advice to researchers who are new to the field and want to implement this method. With careful planning and validation, ATAC-seq can reveal the regulatory landscape of a genome and aid in identifying elements that govern gene expression.

Animals

Comprehensive evaluation of new sequencer T20 and well-established T7 with 507 human samples.

The DNBSEQ-T20×2 (T20) sequencer, developed by MGI Tech, enables cost-effective human whole-genome sequencing (WGS) at 30× coverage for less than $100 per genome. Here, we evaluate the sequencing performance and data quality of the T20 platform by benchmarking it against the established DNBSEQ-T7 (T7) sequencer using 507 samples derived from blood (N = 75), stool (N = 242), and saliva (N = 190). The T20 exhibited lower sequencing quality metrics compared with the T7, with Q20 scores of 95.76%-95.83% and Q30 scores of 87.25%-87.40%, compared with 97.81%-97.93% and 93.26%-93.60%, respectively, for T7 data. Quality differences were more evident toward the end of reads, and PCR-free libraries sequenced on the T20 showed similar reductions in quality scores. The median empirical base error rate estimated from 102 ZymoBIOMICS samples was 0.33%. The T20 demonstrated comparable coverage uniformity to the T7 and showed high concordance in microbiome composition analysis, with a median Bray-Curtis dissimilarity of 0.02. Variant calling performance was highly consistent between the two platforms. Among variants with non-missing genotype calls on both platforms, 94.92% of SNPs and 87.20% of InDels showed concordant genotypes between T20 and T7. Overall, the T20 delivers reliable sequencing accuracy and reproducibility for large-scale genomic and microbiome studies, providing a cost-effective alternative for high-throughput sequencing applications.

Metagenomics

A comparison of mail, telephone, and home interview strategies for household health surveys.

The method of data collection in household health surveys can be a major determinant of cost and data quality. A survey strategy can comprise mail, telephone, or home interview methods, individually or in combination to follow up non-respondents. The purpose of this study in Montreal was to compare cost and data quality of various strategies. Strategies which began with mail or telephone contact, followed by the two other methods, provided response rates as high as a home interview strategy (all between 80 and 90 per cent), for one-half the cost of home interviews when used as the sole method. The telephone response rate was higher than the mail response rate. Comparing different follow-up approaches to strategies beginning with mail or telephone, it proved less costly, and equally effective, to use home interviewing as a last resort for persistent non-respondents. Validity of response (comparing individual responses with records of a government health insurance data bank) and willingness to answer sensitive questions were greatest in mail strategy.

Canada

Additional day-to-day precision estimates based on regional chemistry quality control data.

State-of-the-art precision values are presented for the following serum constituents: aldolase (EC 4.1.2.13), alpha-hydroxybutyrate dehydrogenase (EC 1.1.1.30), cholinesterase (EC 3.1.1.8), cortisol, gamma glutamyl transferase (EC 2.3.2.2), haptoglobin, immunoglobulins, lactic acid, leucine aminopeptidase (EC 3.4.1.1), total lipids, osmolality, protein fractions, T3 uptake, thyroxine and vitamin B12. Precision estimates are based on values reported for four lyophilized serum pools analyzed by participants in the Pennsylvania Association of Clinical Pathologists regional quality control program for clinical chemistry, during 1976, 1977 and 1978. Use of the upper limit of the "most common range" of precision (that range including the 75 percent most precise laboratories) as a warning level for trouble-shooting is advocated.

Blood Chemical Analysis

[Quality of data provided by VESKA medical statistics: the case of the fractured proximal femur].

Within the framework of a retrospective study of the incidence of hip fractures in the canton of Vaud (Switzerland), all cases of hip fracture occurring among the resident population in 1986 and treated in the hospitals of the canton were identified from among five different information sources. Relevant data were then extracted from the medical records. At least two sources of information were used to identify cases in each hospital, among them the statistics of the Swiss Hospital Association (VESKA). These statistics were available for 9 of the 18 hospitals in the canton that participated in the study. The number of cases identified from the VESKA statistics was compared to the total number of cases for each hospital. For the 9 hospitals the number of cases in the VESKA statistics was 407, whereas, after having excluded diagnoses that were actually "status after fracture" and double entries, the total for these hospitals was 392, that is 4% less than the VESKA statistics indicate. It is concluded that the VESKA statistics provide a good approximation of the actual number of cases treated in these hospitals, with a tendency to overestimate this number. In order to use these statistics for calculating incidence figures, however, it is imperative that a greater proportion of all hospitals (50% presently in the canton, 35% nationwide) participate in these statistics.

Data Interpretation, Statistical

Sociocultural context of individual creativity: a transhistorical time-series analysis.

Hypotheses were stated which specify individual creativity as a function of developmental and productive period variables. It was then argued that these hypotheses could be better tested by examining generational fluctuations in creativity. Information from cultural and political archival sources was thus aggregated to form time series spanning 127 generations of European history. Data quality checks, control variables, data transformations, time-lagged comparisons, and trend analyses were used to improve the validity of the causal inferences. While the results varied according to the type of creativity (discursive or presentational) and the degree of achieved eminence, creative development was found to be affected by the following: (a) role model availability, (b) political fragmentation, (c) imperial instability, and (d) political instability.

Achievement

The future of precision oncology and artificial intelligence in Belgium: scenarios and policy responses.

PURPOSE: Precision medicine, also known as personalized medicine, enables the provision of tailored health services to patients. In the prevention, early detection, and treatment of cancers, precision medicine is highly promising, given the increasing use of genomic profiling for diagnosis and adapting therapies in several tumor types. Artificial Intelligence (AI) can support this process by analyzing vast amounts of relevant data. However, high-quality data and financial investments in the health system are essential for the implementation of precision medicine and AI solutions in routine cancer care. DESIGN/METHODOLOGY/APPROACH: Building on the quantitative outcomes of a foresight exercise published in another study, this article collects qualitative data to gain more detailed insights into the future of precision oncology in Belgium and discusses the role of AI in this field. It reports the results of a series of expert workshops, focusing on four hypothetical future scenarios that are centered around technological and economic issues that must be overcome for the widespread use of precision oncology in Belgium. FINDINGS: The study concludes that all four scenarios discussed in the workshops would require supportive policy measures in Belgium, which should go beyond mere technological and economic considerations, such as involving patient associations and the public in policy design or creating multi-disciplinary expert groups for precision medicine. ORIGINALITY/VALUE: To the best of our knowledge, this is the first study to employ foresight methodology to illustrate possible future scenarios, scrutinize feasible approaches for implementing precision oncology in Belgium, and discuss the use of AI in this context.

Belgium

Stability of mean values of organic analytes in lyophilized quality control serum. A study utilizing data from the Quality Assurance Service (QAS) Program of the College of American Pathologists.

The long-term stability of mean values of six organic analytes in 48 commercial pools of lyophilized chemistry quality control serum is evaluated. These pools, provided by five manufacturers, have been used by nine Regional Quality Control Programs participating in the College of American Pathologists Quality Assurance Service. Creatinine yielded stable mean values in 88% of the pools studied. Both albumin and urea nitrogen demonstrated manufacturer and method-related changes in mean values. Bilirubin, cholesterol, and uric acid changes in mean values were independent of control serum manufacturer, predominant in calibrator-standardized automated procedures, and were clustered by years.

Bilirubin