arXiv ScienceSearch

arXiv subjects

Runshi Tang

Publications and source records attributed to Runshi Tang.

9 recordsLinked to original sources

Bandable Cumulant Tensors: Optimal Estimation and Applications in Non-Gaussian Data Modeling

Higher-order cumulants capture the non-Gaussian dependence that covariance misses, but they are hard to use in high dimensions. An order-$d$ cumulant tensor has $p^d$ entries, and the plug-in sample cumulant is generally not even rate-optimal under the tensor spectral norm. For ordered data, both difficulties admit one remedy: assuming that higher-order interactions decay away from the main tensor diagonal, we introduce a bandable cumulant class and a tapered sample cumulant estimator that computes only $O(pk^{d-1})$ local entries at bandwidth $k$ and never forms the full tensor. Under exponential-type tail conditions, we prove nonasymptotic spectral-norm bounds that separate tapering bias from stochastic error, and a minimax lower bound over the same class that matches the leading bias and stochastic terms of the upper bound; for sub-Gaussian observations, the tapered estimator attains the minimax rate whenever $n\gg(k+\log p)^{d-1}$ at the oracle bandwidth $k$, with the ambient dimension entering only through $\log p$. Localization also suppresses the higher-order fluctuations behind this suboptimality, so tapering plays a stronger role here than in bandable covariance estimation. The spectral-norm guarantee transfers directly to downstream tasks, yielding plug-in error bounds for cumulant Yule--Walker estimation in autoregressive models, minimum-distance estimation in moving-average models, and matched-filter source localization in sensor arrays. Simulations corroborate the theory, and real-data analyses of RR-interval, air-quality, and Neuropixels recordings illustrate the resulting stability gains.

stat.ME

Optimal Estimation of Discrete Multiview Distributions under Heteroskedastic Multinomial Sampling

Multiview latent-variable models provide a fundamental framework for discrete data analysis, with applications to latent structure models, topic models, and mixtures of product distributions. In the discrete setting, the joint distribution of the observed views can be represented as a nonnegative low-rank tensor, which we call a multiview density tensor. We study the problem of estimating this tensor from multinomial count data. A key challenge is that multinomial sampling induces heteroskedastic and dependent noise, so the difficulty of estimation depends not only on the ambient dimensions and rank, but also on how the probability mass is distributed across different locations of sample space. We propose a general scaling framework for density tensor estimation under multinomial sampling. This framework leads to a spectral estimator for which we prove a Frobenius-norm upper bound that directly handles heteroskedasticity and negative dependence. For the original multiview model, we obtain fiber-mass-dependent Frobenius upper bounds and minimax lower bounds showing that this dependence is unavoidable. Under $\ell_1$ loss, we develop both oracle and feasible data-driven estimators based on the same scaling principle, establish minimax lower bounds, and show near-optimality for the oracle rule at fixed rank and for slice normalization under bounded slice-to-fiber imbalance. Simulations support the theory and demonstrate the robustness of the proposed methods.

stat.ME

From Reverse Detection--Estimation Gaps to Computational Lower Bounds for Norm Approximation

We develop a general reduction scheme that converts reverse detection--estimation gaps into computational lower bounds for norm approximation. If an efficiently computable estimator has error well below a conjectured computational detection threshold, then a sufficiently accurate norm approximator applied to this estimator would yield an efficient test, transferring average-case detection hardness to worst-case approximation hardness. The ratio between the detection and estimation scales becomes a lower bound on the achievable approximation factor. We apply the reduction in three settings. Using a previously established reverse gap for high-order cumulant tensors, we obtain a conditional lower bound for polynomial-time approximation of the order-$d$ tensor spectral norm. We then establish new reverse gaps for the $k$-sparse matrix spectral norm and a normalized $k$-sparse matrix cut norm under planted-clique hardness. The resulting lower bounds are of order $\sqrt{k}$ up to a logarithmic factor, matching the simple deterministic $\sqrt{k}$ certificates, whereas previous hardness results for the sparse principal component objective rule out only constant factors. Finally, under a model-specific low-degree conjecture, a new reverse gap forces polynomially growing approximation factors for the order-$d$ tensor cut norm when $d\ge3$. These results identify reverse detection--estimation gaps as a systematic source of computational lower bounds for norm approximation.

math.ST

Detection Is Harder Than Estimation in Certain Regimes: Inference for Moment and Cumulant Tensors

We study estimation and detection of high-order moment and cumulant tensors from $n$ i.i.d.\ observations of a $p$-dimensional random vector, with performance measured in tensor spectral norm. Under sub-Gaussianity, we show that the minimax rate for estimating the order-$d$ moment and cumulant tensors is $\sqrt{p/n}\wedge 1$. In contrast to covariance estimation, the sample moment tensor is generally not rate-optimal for $d\ge 3$, and we construct an estimator that attains the minimax rate up to logarithmic factors. On the computational side, we study testing whether the $d$-th order cumulant tensor vanishes after whitening. Using the low-degree polynomial framework, we provide evidence that detection is computationally hard when $n\ll p^{d/2}$. At the same time, we identify a regime in which an efficiently computable estimator achieves error smaller than the separation at which low-degree tests can reliably distinguish the null from the alternative. This reveals an unusual reverse detection--estimation gap: computationally efficient detection can be harder than computationally efficient estimation. The underlying reason is that the relevant loss, tensor spectral norm, is itself NP-hard to compute, creating a new form of computational--statistical gap.

math.ST

A functional tensor model for dynamic multilayer networks with common invariant subspaces and the RKHS estimation

Dynamic multilayer networks are frequently used to describe the structure and temporal evolution of multiple relationships among common entities, with applications in fields such as sociology, economics, and neuroscience. However, exploration of analytical methods for these complex data structures remains limited. We propose a functional tensor-based model for dynamic multilayer networks, with the key feature of capturing the shared structure among common vertices across all layers, while simultaneously accommodating smoothly varying temporal dynamics and layer-specific heterogeneity. The proposed model and its embeddings can be applied to various downstream network inference tasks, including dimensionality reduction, vertex community detection, analysis of network evolution periodicity, visualization of dynamic network evolution patterns, and evaluation of inter-layer similarity. We provide an estimation algorithm based on functional tensor Tucker decomposition and the reproducing kernel Hilbert space framework, with an effective initialization strategy to improve computational efficiency. The estimation procedure can be extended to address more generalized functional tensor problems, as well as to handle missing data or unaligned observations. We validate our method on simulated data and two real-world cases: the dynamic Citi Bike trip network and an international food trade dynamic multilayer network, with each layer corresponding to a different product.

stat.ME

Revisit CP Tensor Decomposition: Statistical Optimality and Fast Convergence

Canonical Polyadic (CP) tensor decomposition is a fundamental technique for analyzing high-dimensional tensor data. While the Alternating Least Squares (ALS) algorithm is widely used for computing CP decomposition due to its simplicity and empirical success, its theoretical foundation, particularly regarding statistical optimality and convergence behavior, remain underdeveloped, especially in noisy, non-orthogonal, and higher-rank settings. In this work, we revisit CP tensor decomposition from a statistical perspective and provide a comprehensive theoretical analysis of ALS under a signal-plus-noise model. We establish non-asymptotic, minimax-optimal error bounds for tensors of general order, dimensions, and rank, assuming suitable initialization. To enable such initialization, we propose Tucker-based Approximation with Simultaneous Diagonalization (TASD), a robust method that improves stability and accuracy in noisy regimes. Combined with ALS, TASD yields a statistically consistent estimator. We further analyze the convergence dynamics of ALS, identifying a two-phase pattern-initial quadratic convergence followed by linear refinement. We further show that in the rank-one setting, ALS with an appropriately chosen initialization attains optimal error within just one or two iterations.

stat.ME

Tensor Decomposition with Unaligned Observations

This paper presents a canonical polyadic (CP) tensor decomposition that addresses unaligned observations. The mode with unaligned observations is represented using functions in a reproducing kernel Hilbert space (RKHS). We introduce a versatile loss function that effectively accounts for various types of data, including binary, integer-valued, and positive-valued types. Additionally, we propose an optimization algorithm for computing tensor decompositions with unaligned observations, along with a stochastic gradient method to enhance computational efficiency. A sketching algorithm is also introduced to further improve efficiency when using the $\ell_2$ loss function. To demonstrate the efficacy of our methods, we provide illustrative examples using both synthetic data and an early childhood human microbiome dataset.

stat.ML

Cocaine Use Prediction with Tensor-based Machine Learning on Multimodal MRI Connectome Data

This paper considers the use of machine learning algorithms for predicting cocaine use based on magnetic resonance imaging (MRI) connectomic data. The study utilized functional MRI (fMRI) and diffusion MRI (dMRI) data collected from 275 individuals, which was then parcellated into 246 regions of interest (ROIs) using the Brainnetome atlas. After data preprocessing, the datasets were transformed into tensor form. We developed a tensor-based unsupervised machine learning algorithm to reduce the size of the data tensor from $275$ (individuals) $\times 2$ (fMRI and dMRI) $\times 246$ (ROIs) $\times 246$ (ROIs) to $275$ (individuals) $\times 2$ (fMRI and dMRI) $\times 6$ (clusters) $\times 6$ (clusters). This was achieved by applying the high-order Lloyd algorithm to group the ROI data into 6 clusters. Features were extracted from the reduced tensor and combined with demographic features (age, gender, race, and HIV status). The resulting dataset was used to train a Catboost model using subsampling and nested cross-validation techniques, which achieved a prediction accuracy of 0.857 for identifying cocaine users. The model was also compared with other models, and the feature importance of the model was presented. Overall, this study highlights the potential for using tensor-based machine learning algorithms to predict cocaine use based on MRI connectomic data and presents a promising approach for identifying individuals at risk of substance abuse.

stat.AP

Mode-wise Principal Subspace Pursuit and Matrix Spiked Covariance Model

This paper introduces a novel framework called Mode-wise Principal Subspace Pursuit (MOP-UP) to extract hidden variations in both the row and column dimensions for matrix data. To enhance the understanding of the framework, we introduce a class of matrix-variate spiked covariance models that serve as inspiration for the development of the MOP-UP algorithm. The MOP-UP algorithm consists of two steps: Average Subspace Capture (ASC) and Alternating Projection (AP). These steps are specifically designed to capture the row-wise and column-wise dimension-reduced subspaces which contain the most informative features of the data. ASC utilizes a novel average projection operator as initialization and achieves exact recovery in the noiseless setting. We analyze the convergence and non-asymptotic error bounds of MOP-UP, introducing a blockwise matrix eigenvalue perturbation bound that proves the desired bound, where classic perturbation bounds fail. The effectiveness and practical merits of the proposed framework are demonstrated through experiments on both simulated and real datasets. Lastly, we discuss generalizations of our approach to higher-order data.

stat.ME