Tuning-Free Covariance-Aware Sequential Shrinkage for Multi-Source Estimation
Modern statistical learning problems often involve multiple related data sets, in which learning efficiency on a target set can be improved by borrowing information from related sources, while source heterogeneity may introduce bias. Existing approaches are limited by suboptimal performance in multi-source settings, insufficient use of covariance information, or the computational burden of tuning procedures. We propose a covariance-aware shrinkage framework that constructs shrinkage directions using covariance information to improve efficiency. We establish finite-sample risk bounds that yield an explicit risk-improving interval for the shrinkage size, making the transfer strength fully data-driven. When multiple source sets are available, we further propose a sequential algorithm that applies source-specific shrinkage updates according to a priority score. Under a mixed-source separation condition, the estimator asymptotically attains the oracle risk and, when the pooled discrepancy is large, strictly improves over single-step shrinkage. The framework is further extended to a broad class of smooth M-estimation problems. Numerical studies show substantial gains over competing methods, especially when the source data sets are highly heterogeneous.