Search PubMed⌕ Search

PubMed · 16321746

Perceptually based FROC analysis.

Abstract

RATIONALE AND OBJECTIVES: Analysis of reading data when cases have multiple targets and/or the reader is required to localize targets is difficult. One approach to this free-response operating characteristic (FROC) problem is for images to be segmented (eg, with quadrants) by the investigator and a segment-level analysis be conducted with the case as a nesting factor. In this report, we introduce an alternative method that uses the visual scan path of the reader to segment the image. We evaluate the new method by applying it to data from a mammography reading experiment. MATERIALS AND METHODS: The gaze scan path of one radiologist was recorded as she scanned 40 mammograms for masses and microcalcifications. The observer is an experienced mammographer and was not one of the authors. In addition, the reader provided a rating indicating the degree of suspicion for any suspected targets she identified and localized. We then established "perceptual regions" by using a clustering algorithm on the visual fixations. We combined ratings given to specific locations indicated by the reader with the segmentation from the visual scan to generate a series of ratings classified for whether the perceptually based region associated with the rating contained or did not contain a known target. We analyzed data generated by our method from all 40 cases by using the conventional maximum-likelihood method based on the binormal model. Finally, we tested goodness-of-fit of the binormal model to the data by using chi-square. RESULTS: Maximum-likelihood estimation led to a model that did not fit the data (P < .001). However, examination of the observed and expected counts suggests that the binormal assumption does not hold for segments that contain targets and a bimodal distribution model might be preferred. CONCLUSION: Our new method provides an alternative approach to analysis of the FROC experiment. It needs to be developed further. Specifically, we propose that a mixture model extension of the binormal model be developed for ratings data arising from perceptually based FROC experiments. A disadvantage to our method is the requirement to record the scan path of the reader. However, we believe that adding such information to receiver operating characteristic (ROC) curve analysis will pay off when appropriate statistical models have been identified because we believe our data support our hypothesis that the perceptual scanning of images by humans deconvolves interpretation correlation. If true, this hypothesis implies that conventional statistical methods for ROC analysis based on independent data can be applied to the analysis of FROC data after conditioning on the scan path of the observer.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Rachna Arora, Harold L Kundel, Craig A Beam. 2005. Perceptually based FROC analysis.. https://doi.org/10.1016/j.acra.2005.06.015

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

A comparison of regression trees, logistic regression, generalized additive models, and multivariate adaptive regression splines for predicting AMI mortality.

Clinicians and health service researchers are frequently interested in predicting patient-specific probabilities of adverse events (e.g. death, disease recurrence, post-operative complications, hospital readmission). There is an increasing interest in the use of classification and regression trees (CART) for predicting outcomes in clinical studies. We compared the predictive accuracy of logistic regression with that of regression trees for predicting mortality after hospitalization with an acute myocardial infarction (AMI). We also examined the predictive ability of two other types of data-driven models: generalized additive models (GAMs) and multivariate adaptive regression splines (MARS). We used data on 9484 patients admitted to hospital with an AMI in Ontario. We used repeated split-sample validation: the data were randomly divided into derivation and validation samples. Predictive models were estimated using the derivation sample and the predictive accuracy of the resultant model was assessed using the area under the receiver operating characteristic (ROC) curve in the validation sample. This process was repeated 1000 times-the initial data set was randomly divided into derivation and validation samples 1000 times, and the predictive accuracy of each method was assessed each time. The mean ROC curve area for the regression tree models in the 1000 derivation samples was 0.762, while the mean ROC curve area of a simple logistic regression model was 0.845. The mean ROC curve areas for the other methods ranged from a low of 0.831 to a high of 0.851. Our study shows that regression trees do not perform as well as logistic regression for predicting mortality following AMI. However, the logistic regression model had performance comparable to that of more flexible, data-driven models such as GAMs and MARS.

Data Interpretation, Statistical↗

Multiple linear regression with some correlated errors: classical and robust methods.

In this paper we consider classical and robust methods of estimation and diagnostics for the multiple linear regression model when some of the errors are correlated. This work was motivated by the analysis of a medical data set, from an observational study aimed at identifying factors affecting the outcome of a surgical method for the correction of scoliosis (abnormal lateral spinal curvature). There are 392 observations but some of them are on the same patient (double curves). It seems adequate to consider a multiple linear regression model but, since it is not desirable to discard the double curves, the assumption of non-correlated errors is clearly violated, and this is indeed confirmed by related diagnostics on the residuals (Durbin-Watson test). A more appropriate model retains the linear structure but allows for non-null correlation between the errors on the same patient. We propose two different procedures for the estimation of the parameters of the linear model and the correlation parameters: maximum likelihood assuming normal errors and a robustified version obtained by plugging-in results from robust linear regression. The latter procedure is designed to be resistant to outlying observations or error distributions with heavy tails and has produced the most satisfactory results for the analysed data set.

Data Interpretation, Statistical↗

Adaptive design method based on sum of p-values.

Bauer and Kohne proposed an adaptive design using Fisher's combination of independent p-values based on subsamples from different stages (Biometrics 1994; 50(4):1029-1041). Their method provides great flexibility in the selection of statistical methods for hypothesis testing of subsamples. However, the choices for the stopping boundaries are not flexible enough to meet practical needs (Biometrics 2001; 57(3): 886-891). In this paper, an adaptive design method is proposed using linear combination of the independent p-values. The method provides great flexibility in the selection of stopping boundaries and no numerical integration is required for the two-stage designs. The stopping boundaries and p-values can be calculated manually. The operating characteristics of the adaptive designs are studied using computer simulations with and without sample size adjustment. Examples are presented for superiority and non-inferiority trials with different endpoints (normal, binary, and survival) under different adaptations. The statistical efficiency of the proposed method is compared with other methods based on conditional power.

Data Interpretation, Statistical↗