arXiv ScienceSearch

arXiv subjects

Orestis Efthimiou

Publications and source records attributed to Orestis Efthimiou.

2 recordsLinked to original sources

Estimating heterogeneous treatment effects from randomised trials: a comparison of the risk modelling and treatment effect modelling approaches

Introduction Risk modelling (RM) and treatment effect modelling (EM) are two approaches to build models to estimate heterogeneous treatment effects from randomized control trials (RCT). RM is a two-stage approach; it estimates a predicted baseline risk at the first stage and then includes it as the only effect modifier in a regression model in the second stage. EM is a full treatment interaction multivariable model. Both approaches have theoretical advantages and limitations, but a thorough comparison including a simulation study is missing. Methods In a theoretical review of the two approaches, we present their underlying assumptions and theoretical advantages and disadvantages. Based on our theoretical review, we design simulation scenarios for RCTs with a dichotomous outcome and evaluate the performance of both approaches with respect to the root mean square error and the bias in the predicted risk difference. Results In the theoretical part we argue that baseline risk is a treatment effect modifier in many clinical situations. RM is a dimensionality reduction approach which, however, makes strong assumptions about the role of prognostic factors modifying the treatment effect. EM's greater flexibility is a possible advantage when the sample size is large. In most simulated scenarios RM performs better than EM, even when the assumptions underlying RM are not fully met. The advantage of RM diminishes as sample size increases. Conclusion When choosing between RM and EM the available sample size and the plausibility of their underlying assumptions should be considered.

stat.ME

Multiple imputation of incomplete multilevel data using Heckman selection models

Missing data is a common problem in medical research, and is commonly addressed using multiple imputation. Although traditional imputation methods allow for valid statistical inference when data are missing at random (MAR), their implementation is problematic when the presence of missingness depends on unobserved variables, i.e. the data are missing not at random (MNAR). Unfortunately, this MNAR situation is rather common, in observational studies, registries and other sources of real-world data. While several imputation methods have been proposed for addressing individual studies when data are MNAR, their application and validity in large datasets with multilevel structure remains unclear. We therefore explored the consequence of MNAR data in hierarchical data in-depth, and proposed a novel multilevel imputation method for common missing patterns in clustered datasets. This method is based on the principles of Heckman selection models and adopts a two-stage meta-analysis approach to impute binary and continuous variables that may be outcomes or predictors and that are systematically or sporadically missing. After evaluating the proposed imputation model in simulated scenarios, we illustrate it use in a cross-sectional community survey to estimate the prevalence of malaria parasitemia in children aged 2-10 years in five subregions in Uganda.

stat.ME