arXiv Science⌕ Search

arXiv · 2610.08443

Scalable Regularized Vector Multiplicative Error Models for Positive-valued Financial Time Series

Abstract

The logarithmic multiplicative error model (log-vMEM) has been useful in modeling and forecasting multivariate positive-valued financial time series. The number of parameters grow rapidly with the dimension of the system and the lag order, making estimation computationally demanding in high-dimensional settings. This paper describes regularized estimation via hierarchical lag structures (Nicholson et al., 2020) for log-vMEM models with multivariate gamma error distribution of Tsionas (2004). The parameter estimation is performed using a blockwise coordinate descent algorithm with a Gauss-Seidel-style update scheme (Wright, 2015). This enables an efficient computation strategy compared to traditional penalized maximum likelihood approaches. The competing models are juxtaposed against each other by combining three hierarchical lag structures (componentwise, elementwise, own-other) and four penalties(group-lasso, adaptive group-lasso, group-mcp, and group-scad). Extensive simulation runs have been performed to test the parameter recovery for both the unpenalized and the penalized models. We apply the proposed methods to model the joint dynamics of robust intraday realized volatility measures for Microsoft (NASDAQ: MSFT) for the competing models. The numerical integration step of the log-likelihood is identified to be the principal computational bottleneck. We address this issue by using GPU-accelerated quadrature integration thus improving computational scalability of the proposed models.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Rohan Hemant Chhatre, Chiranjit Dutta, Nalini Ravishanker, Sumanta Basu. 2026-10-06. Scalable Regularized Vector Multiplicative Error Models for Positive-valued Financial Time Series. https://arxiv.org/abs/2610.08443

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Two-Loop Stochastic Mirror Langevin Algorithms for Constrained Sampling

We study the problem of sampling from a target distribution $π(x)\propto e^{-f(x)}$ supported on a convex set $ X\subseteq\mathbb R^d$, when the potential $f$ is accessible only through a stochastic first-order oracle. Mirror Langevin algorithms provide a natural approach to constrained sampling by transporting the problem to an unconstrained dual space and discretizing the resulting Mirror Langevin diffusion. Existing implementations, however, typically use a fixed discretization step size and consequently retain a nonvanishing discretization bias at any fixed step size. Moreover, their direct extension to settings with noisy gradient information entails the challenge of controlling both discretization and stochastic-oracle error. We study a stochastic first-order version of the Mirror Langevin Algorithm (sFO-MLA) and, as our main contribution, develop a warm-started two-loop implementation in which an outer loop progressively decreases the step size while an inner loop runs sFO-MLA (with a fixed step size) for an appropriately chosen epoch length. The construction provides a principled schedule linking step sizes and epoch lengths, so that successive epochs warm-start from increasingly accurate distributions rather than repeatedly paying the cost of mixing from a cold start. We establish finite-time Wasserstein guarantees for sFO-MLA that explicitly separate mixing, Euler--Maruyama discretization, and stochastic-gradient errors. These bounds yield a fixed horizon rate of $\widetilde O(T^{-1/2})$ and show that the two-loop scheme removes the associated logarithmic penalty, attaining the canonical $O(T^{-1/2})$ rate under a geometric step-size schedule and corresponding epoch lengths. We illustrate the methodology in two statistically distinct settings.

stat.CO↗

BKP: An R Package for Beta Kernel Process Modeling

Estimating input-dependent probability surfaces from binary, binomial, categorical, or multinomial response data is a common task in statistics and machine learning. Latent Gaussian process classifiers provide flexible nonparametric models for such problems, but posterior inference with discrete responses typically requires approximation or simulation. We discuss an implementation of probability-scale beta and Dirichlet kernel models in the \pkg{BKP} package for \proglang{R}. The package implements the Beta Kernel Process (BKP), which uses kernel-weighted pseudo-count aggregation and beta-binomial conjugacy to obtain closed-form conjugate posterior summaries and posterior predictive distributions for binomial probabilities. It also implements the Dirichlet Kernel Process (DKP) for multi-class responses, together with TwinBKP and TwinDKP, scalable twinning-based global-local approximations for larger datasets. The resulting workflow supports transparent kernel-weighted evidence borrowing, several kernel families, fixed and data-adaptive priors, effective-sample-size calibration, loss-based hyperparameter tuning, and standard S3 methods for fitting, prediction, simulation, visualization, and extraction of posterior summaries. Reproducible examples demonstrate probability-surface estimation, binary and multi-class classification, computational comparison, and real-data applications to \emph{Loa loa} infection prevalence mapping and Mourning Warbler distribution modeling.

stat.CO↗

High-dimensional reliability-based design optimization using stochastic emulators

Reliability-based design optimization (RBDO) is traditionally formulated as a nested optimization and reliability problem and remains computationally demanding, especially in high dimensions. This paper proposes a new RBDO framework based on a stochastic simulator viewpoint, in which the deterministic limit-state function and uncertain model inputs are combined into a unified stochastic representation. For a given design, the system response is characterized directly through its conditional output distribution rather than through an explicit limit-state function. Stochastic emulators are constructed in the design space to approximate this conditional distribution, enabling semi-analytical evaluation of failure probabilities or associated quantiles without Monte Carlo simulation. Two approaches are considered: generalized lambda models (GLaM) and stochastic polynomial chaos expansions (SPCE). Both yield deterministic mappings between design variables and reliability constraints, thereby eliminating the classical double-loop structure and enabling standard deterministic optimization. The approach is assessed on benchmark problems ranging from low to very high dimension, including stochastic excitation, and compared with Kriging in the full input space and heteroscedastic Gaussian processes in the design space. The proposed method provides substantial computational gains, particularly for high-dimensional random inputs, while retaining comparable efficiency to Kriging in low dimensions. Unlike heteroscedastic Gaussian processes, GLaM and SPCE also avoid restrictive assumptions on the conditional response distribution.

stat.CO↗