Search PubMed⌕ Search

PubMed · 10544307

Hazard function estimators: a simulation study.

Abstract

Kernel-based methods for the smooth, non-parametric estimation of the hazard function have received considerable attention in the statistical literature. Although the mathematical properties of the kernel-based hazard estimators have been carefully studied, their statistical properties have not. We reviewed various kernel-based methods for hazard function estimation from right-censored data and compared the statistical properties of these estimators through computer simulations. Our simulations covered seven distributions, three levels of random censoring, four types of bandwidth functions, two sample sizes and three types of boundary correction. We conducted a total of 504 simulation experiments with 500 independent samples each. Our results confirmed the advantages of two recent innovations in kernel estimation - boundary correction and locally optimal bandwidths. The median relative improvement (decrease) in mean square error over fixed-bandwidth estimators without boundary correction was 3 per cent for fixed-bandwidth estimators with left boundary correction, 52 per cent locally optimal bandwidths without boundary correction, and 66 per cent for locally optimal bandwidths with left boundary correction. The locally optimal bandwidth estimators with left boundary correction also outperformed three previously published and publicly available algorithms, with median relative improvements in mean square error of 31 per cent, 77 per cent and 80 per cent.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

K R Hess, D M Serachitopol, B W Brown. 1999-11-30. Hazard function estimators: a simulation study.. https://doi.org/10.1002/(sici)1097-0258(19991130)18%3A22%3C3075%3A%3Aaid-sim244%3E3.0.co%3B2-6

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related citations

Generating correlated data for omics simulation.

Simulation of realistic omics data is a key input for benchmarking studies that help users obtain optimal computational pipelines. Omics data involves large numbers of measured features on each sample and these measures are generally correlated with each other. However, simulation too often ignores these correlations, perhaps due to computational and statistical hurdles of doing so. To alleviate this, we describe three approaches for generating omics-scale data with correlated measures which mimic real datasets. These approaches are all based on a Gaussian copula approach with a covariance matrix that decomposes into a diagonal part and a low-rank part. This decomposition allows for extremely efficient simulation, overcoming a hurdle for adoption of past methods. We use these approaches to demonstrate the importance of including correlation in two benchmarking applications. First, we show that variance of results from the popular DESeq2 method increases when dependence is included. Second, we demonstrate that CYCLOPS, a method for inferring circadian time of collection from transcriptomics, improves in performance when given gene-gene dependencies in some circumstances. We provide an R package, dependentsimr, that has efficient implementations of these methods and can generate dependent data with arbitrary marginal distributions, including discrete (binary, ordered categorical, Poisson, negative binomial), continuous (normal), or with an empirical distribution.

Computer Simulation↗

Addressing current challenges in cancer immunotherapy with mathematical and computational modelling.

The goal of cancer immunotherapy is to boost a patient's immune response to a tumour. Yet, the design of an effective immunotherapy is complicated by various factors, including a potentially immunosuppressive tumour microenvironment, immune-modulating effects of conventional treatments and therapy-related toxicities. These complexities can be incorporated into mathematical and computational models of cancer immunotherapy that can then be used to aid in rational therapy design. In this review, we survey modelling approaches under the umbrella of the major challenges facing immunotherapy development, which encompass tumour classification, optimal treatment scheduling and combination therapy design. Although overlapping, each challenge has presented unique opportunities for modellers to make contributions using analytical and numerical analysis of model outcomes, as well as optimization algorithms. We discuss several examples of models that have grown in complexity as more biological information has become available, showcasing how model development is a dynamic process interlinked with the rapid advances in tumour-immune biology. We conclude the review with recommendations for modellers both with respect to methodology and biological direction that might help keep modellers at the forefront of cancer immunotherapy development.

Computer Simulation↗

A generalized concordance correlation coefficient for continuous and categorical data.

This paper discusses a generalized version of the concordance correlation coefficient for agreement data. The concordance correlation coefficient evaluates the accuracy and precision between two measures, and is based on the expected value of the squared function of distance. We have generalized this coefficient by applying alternative functions of distance to produce more robust versions of the concordance correlation coefficient. In this paper we extend the application of this class of estimators to categorical data as well, and demonstrate similarities to the kappa and weighted kappa statistics. We also introduce a stratified concordance correlation coefficient which adjusts for explanatory factors, and an extended concordance correlation coefficient which measures agreement among more than two responses. With these extensions, the generalized concordance correlation coefficient provides a unifying approach to assessing agreement among two or more measures that are either continuous or categorical in scale.

Computer Simulation↗