arXiv ScienceSearch

arXiv · 2512.02037

Statistical Arbitrage in Polish Equities Market Using Deep Learning Techniques

Abstract

We study a systematic approach to a popular Statistical Arbitrage technique: Pairs Trading. Instead of relying on two highly correlated assets, we replace the second asset with a replication of the first using risk factor representations. These factors are obtained through Principal Components Analysis (PCA), exchange traded funds (ETFs), and, as our main contribution, Long Short Term Memory networks (LSTMs). Residuals between the main asset and its replication are examined for mean reversion properties, and trading signals are generated for sufficiently fast mean reverting portfolios. Beyond introducing a deep learning based replication method, we adapt the framework of Avellaneda and Lee (2008) to the Polish market. Accordingly, components of WIG20, mWIG40, and selected sector indices replace the original S&P500 universe, and market parameters such as the risk free rate and transaction costs are updated to reflect local conditions. We outline the full strategy pipeline: risk factor construction, residual modeling via the Ornstein Uhlenbeck process, and signal generation. Each replication technique is described together with its practical implementation. Strategy performance is evaluated over two periods: 2017-2019 and the recessive year 2020. All methods yield profits in 2017-2019, with PCA achieving roughly 20 percent cumulative return and an annualized Sharpe ratio of up to 2.63. Despite multiple adaptations, our conclusions remain consistent with those of the original paper. During the COVID-19 recession, only the ETF based approach remains profitable (about 5 percent annual return), while PCA and LSTM methods underperform. LSTM results, although negative, are promising and indicate potential for future optimization.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Marek Adamczyk, Michał Dąbrowski. 2025-11-20. Statistical Arbitrage in Polish Equities Market Using Deep Learning Techniques. https://arxiv.org/abs/2512.02037

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Does Crypto Sentiment Extremity Widen Estimated Spreads? Evidence Depends on the Specification

We examine whether extreme values of the Crypto Fear & Greed Index are associated with a daily high-low spread estimate for Bitcoin. The sample contains 2,896 BTC/USDT observations from February 2018 to January 2026. We find an unconditional extreme-minus-neutral gap of 61.99 basis points. After close-to-close realised-volatility-quintile demeaning it is 24.79 basis points, although none of the five separate quintile contrasts survives Holm correction. With quadratic realised-volatility and strictly lagged momentum controls, the HAC estimate is 11.81 basis points (95% CI [-2.31,25.93], p=.101). A fixed non-parametric stratification gives 20.44 basis points (p=.0195 under circular shifts), while separate models for a zero-floored estimate's incidence and positive magnitude are imprecise. The results therefore show only a descriptive, specification-dependent association. We conclude that they do not establish a stable or causal liquidity premium.

q-fin.ST

Do Cryptocurrency Markets Differentiate Infrastructure from Regulatory Shocks? A Multi-Moment Event Study with Dependence-Robust Inference

Do cryptocurrency markets respond differently to infrastructure and regulatory shocks? We study returns and conditional variance on a shared sample of 50 events and six assets (January 2019--August 2025), using GJR-GARCH-X models and dependence-aware inference. Treating event inclusion as a design parameter, we trace the variance differential across inclusion screens. Curated high-salience events yield a $3.49\times$ point-estimate multiplier, whereas a mechanical impact filter on a broad reconstructed candidate pool yields approximately $0.5$--$1.6\times$. This pattern is descriptive and selection-conditional, not an inferential comparison between screens. The curated variance differential is not significant against its fitted sharp per-asset-equality null under the conditional fixed-path Student-$t$-copula bootstrap ($p\approx0.39$). Floored recursive sensitivity gives one-sided $p=0.025$ at baseline and $0.041$ with asset-specific high-variance-regime controls, so the verdict depends on inference scheme and implementation. Across reported dependence inputs, the six-contrast effective sample size is $1.33$--$2.35$; one-sided effective-df sensitivity gives $p=0.044$--$0.116$. Earlier significance from treating correlated per-asset coefficients as independent samples is not robust to dependence and heavy-tail corrections. The cumulative-abnormal-return difference is $+8.69$ percentage points (event-level block-bootstrap $p=0.202$). The asymmetry remains directional, selection-conditional and unresolved. The contribution is an inference ladder and an internal Monte-Carlo calibration study, demonstrated through correction of the author's earlier significance claim.

q-fin.ST

Modeling financial time series with $ϕ^{4}$ quantum field theory

We use a $ϕ^{4}$ quantum field theory with inhomogeneous couplings and explicit symmetry-breaking to model an ensemble of financial time series from the S$\&$P 500 index. The continuum nature of the $ϕ^4$ theory avoids the inaccuracies that occur in Ising-based models which require a discretization of the time series. We demonstrate this using the example of the 2008 global financial crisis. The $ϕ^{4}$ quantum field theory is expressive enough to reproduce the higher-order statistics such as the market kurtosis, which can serve as an indicator of possible market shocks. Accurate reproduction of high kurtosis is absent in binarized models. Therefore Ising models, despite being widely employed in econophysics, are incapable of fully representing empirical financial data, a limitation not present in the generalization of the $ϕ^{4}$ scalar field theory. We then investigate the scaling properties of the $ϕ^{4}$ machine learning algorithm and extract exponents which govern the behavior of the learned couplings (or weights and biases in ML language) in relation to the number of stocks in the model. Finally, we use our model to forecast the price changes of the AAPL, MSFT, and NVDA stocks. We conclude by discussing how the $ϕ^{4}$ scalar field theory could be used to build investment strategies and the possible intuitions that the QFT operations of dimensional compactification and renormalization can provide for financial modelling.

q-fin.ST