Search PubMed⌕ Search

PubMed · 11939058

[Identify influential points].

Abstract

Influential points are the points with excessive influence on result and detected by influential function theta-theta(i). theta-theta(i) is a multidimensional vector. For convenience, people always consider some quantitative function g (theta-theta(i)) to measure the vector theta-theta(i), but how to define g is a difficult problem. This paper proposes a new quantitative diagnostic statistics for identifying influential points for all parametrical statistical analysis. The method is more appropriate than cook's method for regression analysis. In this study an example is used to illustrate the procedure in COX proportional hazard model.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

H Lin, Y Su. 1998-09-30. [Identify influential points].. https://pubmed.ncbi.nlm.nih.gov/11939058/

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

SAS and SPLUS programs to perform Cox regression without convergence problems.

When analyzing survival data, the parameter estimates and consequently the relative risk estimates of a Cox model sometimes do not converge to finite values. This phenomenon is due to special conditions in a data set and is known as 'monotone likelihood'. Statistical software packages for Cox regression using the maximum likelihood method cannot appropriately deal with this problem. A new procedure to solve the problem has been proposed by G. Heinze, M. Schemper, A solution to the problem of monotone likelihood in Cox regression, Biometrics 57 (2001). It has been shown that unlike the standard maximum likelihood method, this method always leads to finite parameter estimates. We developed a SAS macro and an SPLUS library to make this method available from within one of these widely used statistical software packages. Our programs are also capable of performing interval estimation based on profile penalized log likelihood (PPL) and of plotting the PPL function as was suggested by G. Heinze, M. Schemper, A solution to the problem of monotone likelihood in Cox regression, Biometrics 57 (2001).

Proportional Hazards Models↗

SAR modeling of unbalanced data sets.

The increased acceptance of SAR approaches to hazard identification has led us to investigate methods to improve the predictive performance of SAR models. In the present study we demonstrate that although on theoretical grounds the ratio of active to inactive chemicals in the learning set should be unity, SAR models can "tolerate" an unbalanced range in ratios from 3:1 (i.e., 75% actives) to 1:2 (i.e., 33% actives) and still perform adequately. On the other hand SAR models derived from learning sets with ratios in excess of 4:1 (80% actives), even when corrected for the initial ratio do not perform satisfactorily.

Proportional Hazards Models↗

The accelerated failure time model: a useful alternative to the Cox regression model in survival analysis.

For the past two decades the Cox proportional hazards model has been used extensively to examine the covariate effects on the hazard function for the failure time variable. On the other hand, the accelerated failure time model, which simply regresses the logarithm of the survival time over the covariates, has seldom been utilized in the analysis of censored survival data. In this article, we review some newly developed linear regression methods for analysing failure time observations. These procedures have sound theoretical justification and can be implemented with an efficient numerical method. The accelerated failure time model has an intuitive physical interpretation and would be a useful alternative to the Cox model in survival analysis.

Proportional Hazards Models↗