arXiv ScienceSearch

arXiv subjects

Claudio Busatto

Publications and source records attributed to Claudio Busatto.

3 recordsLinked to original sources

Fast QR updating methods for statistical applications

This paper introduces fast R updating algorithms specifically designed for statistical applications, including regression, filtering, and model selection, where data structures change frequently. Although traditional QR decomposition is essential for matrix operations, it becomes computationally intensive when dynamically updating the design matrix in statistical models. The proposed algorithms efficiently update the R matrix without the need for recalculation of Q, thereby significantly reducing computational costs in practical computational scenarios. The provision of scalable solutions for high-dimensional regression models is a key strength of these algorithms, enhancing the feasibility of large-scale statistical analyses and model selection in data-intensive fields. A thorough simulation study and the analysis of real-world data demonstrate that the methods achieve a substantial reduction in computational time without compromising accuracy. The discussion illustrates the benefits of these algorithms across a wide range of models and applications in statistics and machine learning.

stat.ME

Informative co-data learning for high-dimensional Horseshoe regression

High-dimensional data often arise from clinical genomics research to infer relevant predictors of a particular trait. A way to improve the predictive performance is to include information on the predictors derived from prior knowledge or previous studies. Such information is also referred to as ``co-data''. To this aim, we develop a novel Bayesian model for including co-data in a high-dimensional regression framework, called Informative Horseshoe regression (infHS). The proposed approach regresses the prior variances of the regression parameters on the co-data variables, improving variable selection and prediction. We implement both a Gibbs sampler and a Variational approximation algorithm. The former is suited for applications of moderate dimensions which, besides prediction, target posterior inference, whereas the computational efficiency of the latter allows handling a very large number of variables. We show the benefits from including co-data with a simulation study. Eventually, we demonstrate that infHS outperforms competing approaches for two genomics applications.

stat.ME

Inference of multiple high-dimensional networks with the Graphical Horseshoe prior

We develop a novel full-Bayesian approach for multiple correlated precision matrices, called multiple Graphical Horseshoe (mGHS). The proposed approach relies on a novel multivariate shrinkage prior based on the Horseshoe prior that borrows strength and shares sparsity patterns across groups, improving posterior edge selection when the precision matrices are similar. On the other hand, there is no loss of performance when the groups are independent. Moreover, mGHS provides a similarity matrix estimate, useful for understanding network similarities across groups. We implement an efficient Metropolis-within-Gibbs for posterior inference; specifically, local variance parameters are updated via a novel and efficient modified rejection sampling algorithm that samples from a three-parameter Gamma distribution. The method scales well with respect to the number of variables and provides one of the fastest full-Bayesian approaches for the estimation of multiple precision matrices. Finally, edge selection is performed with a novel approach based on model cuts. We empirically demonstrate that mGHS outperforms competing approaches through both simulation studies and an application to a bike-sharing dataset.

stat.ME