Search PubMed⌕ Search

PubMed · 16342180

Genotyping errors, pedigree errors, and missing data.

Abstract

Our group studied the effects of genotyping errors, pedigree errors, and missing data on a wide range of techniques, with a focus on the role of single-nucleotide polymorphisms (SNPs). Half of our group used simulated data, and half of our group used data from the Collaborative Study on the Genetics of Alcoholism (COGA). The simulated data had no missing genotypes and no genotyping errors, so our group, as a whole, removed data and introduced artificial errors to study the robustness of various techniques. Our teams showed that genotyping errors are less detectable and may have a greater impact on SNPs than on microsatellites, but recently developed methods that account for genotyping errors help reduce false positives, and the assumptions of these methods appear to be supported by observations from repeated genotyping. The ability to detect linkage disequilibrium (LD) was also substantially reduced by missing data; this in turn could affect tagging SNPs chosen to generate haplotypes. In the COGA sample, genotyping measurements were repeated in three ways. First, full-genome screens were performed on three sets of markers: 328 microsatellites, 11,560 SNPs from the Affymetrix GeneChip Mapping 10 K Array marker set, and 4,720 SNPs from the Illumina Linkage III panel. Second, the entire Affymetrix marker set was typed on the same 184 individuals by two different laboratories. Finally, the Affymetrix and Illumina marker panels had 94 SNPs in common. Our teams showed that both SNPs and microsatellites can be readily used to identify pedigree errors, and that SNPs have fewer genotyping errors and a low inconsistency rate. However, a fairly high rate of no-calls, especially for the Affymetrix platform, suggests that the inconsistency rate may be higher than observed.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Anthony L Hinrichs, Brian K Suarez. 2005. Genotyping errors, pedigree errors, and missing data.. https://doi.org/10.1002/gepi.20120

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Pooled association genome scanning for alcohol dependence using 104,268 SNPs: validation and use to identify alcoholism vulnerability loci in unrelated individuals from the collaborative study on the genetics of alcoholism.

Association genome scanning can identify markers for the allelic variants that contribute to vulnerability to complex disorders, including alcohol dependence. To improve the power and feasibility of this approach, we report validation of "100k" microarray-based allelic frequency assessments in pooled DNA samples. We then use this approach with unrelated alcohol-dependent versus control individuals sampled from pedigrees collected by the Collaborative Study on the Genetics of Alcoholism (COGA). Allele frequency differences between alcohol-dependent and control individuals are assessed in quadruplicate at 104,268 autosomal SNPs in pooled samples. One hundred eighty-eight SNPs provide (1) the largest allele frequency differences between dependent versus control individuals; (2) t values >or= 3 for these differences; and (3) clustering, so that 51 relatively small chromosomal regions contain at least three SNPs that satisfy criteria 1 and 2 above (Monte Carlo P = 0.00034). These positive SNP clusters nominate interesting genes whose products are implicated in cellular signaling, gene regulation, development, "cell adhesion," and Mendelian disorders. The results converge with linkage and association results for alcohol and other addictive phenotypes. The data support polygenic contributions to vulnerability to alcohol dependence. These SNPs provide new tools to aid the understanding, prevention, and treatment of alcohol abuse and dependence.

Alcoholism↗

Determining a cut-off on the Severity of Dependence Scale (SDS) for alcohol dependence.

Optimal cut-off points on the Severity of Dependence Scale (SDS) indicative of clinically significant dependence have been determined for a range of substance types. This study aims to determine a cut-off point on SDS that discriminates between the presence and absence of a DSM-IV diagnosis of alcohol dependence. A structured interview was administered to 90 alcohol users in Sydney, Australia. Receiver Operating Characteristic curve analysis confirmed the utility of the SDS-alcohol for characterising and diagnosing persons with respect to their alcohol-dependent status to an accuracy of 85%. A SDS score of 3 or above was determined as optimal for characterising alcohol dependence. Evidence is also provided confirming that the SDS-alcohol is a valid, reliable uni-dimensional scale for measuring alcohol dependence. It has been demonstrated that the SDS-alcohol can be used to characterise an individual's alcohol-dependent status. A cut-off value for SDS-alcohol provides additional meaning and value to the scale for clients and clinicians and will enable researchers to characterise the prevalence of alcohol dependence in their target populations.

Alcoholism↗