arXiv ScienceSearch

arXiv subjects

Matteo Rizzato

Publications and source records attributed to Matteo Rizzato.

5 recordsLinked to original sources

Generative Adversarial Networks Applied to Synthetic Financial Scenarios Generation

The finance industry is producing an increasing amount of datasets that investment professionals can consider to be influential on the price of financial assets. These datasets were initially mainly limited to exchange data, namely price, capitalization and volume. Their coverage has now considerably expanded to include, for example, macroeconomic data, supply and demand of commodities, balance sheet data and more recently extra-financial data such as ESG scores. This broadening of the factors retained as influential constitutes a serious challenge for statistical modeling. Indeed, the instability of the correlations between these factors makes it practically impossible to identify the joint laws needed to construct scenarios. Fortunately, spectacular advances in Deep Learning field in recent years have given rise to GANs. GANs are a type of generative machine learning models that produce new data samples with the same characteristics as a training data distribution in an unsupervised way, avoiding data assumptions and human induced biases. In this work, we are exploring the use of GANs for synthetic financial scenarios generation. This pilot study is the result of a collaboration between Fujitsu and Advestis and it will be followed by a thorough exploration of the use cases that can benefit from the proposed solution. We propose a GANs-based algorithm that allows the replication of multivariate data representing several properties (including, but not limited to, price, market capitalization, ESG score, controversy score,. . .) of a set of stocks. This approach differs from examples in the financial literature, which are mainly focused on the reproduction of temporal asset price scenarios. We also propose several metrics to evaluate the quality of the data generated by the GANs. This approach is well fit for the generation of scenarios, the time direction simply arising as a subsequent (eventually conditioned) generation of data points drawn from the learned distribution. Our method will allow to simulate high dimensional scenarios (compared to $\lesssim10$ features currently employed in most recent use cases) where network complexity is reduced thanks to a wisely performed feature engineering and selection. Complete results will be presented in a forthcoming study.

q-fin.CP

Extremely expensive likelihoods: A variational-Bayes solution for precision cosmology

We present a variational-Bayes solution to compute non-Gaussian posteriors from extremely expensive likelihoods. Our approach is an alternative for parameter inference when MCMC sampling is numerically prohibitive or conceptually unfeasible. For example, when either the likelihood or the theoretical model cannot be evaluated at arbitrary parameter values, but only previously selected values, then traditional MCMC sampling is impossible, whereas our variational-Bayes solution still succeeds in estimating the full posterior. In cosmology, this occurs e.g. when the parametric model is based on costly simulations that were run for previously selected input parameters. We demonstrate the applicability of our posterior construction on the KiDS-450 weak lensing analysis, where we reconstruct the original KiDS MCMC posterior at 0.6% of its former numerical posterior evaluations. The reduction in numerical cost implies that systematic effects which formerly exhausted the numerical budget could now be included.

astro-ph.CO

Breaking degeneracies with the Sunyaev-Zeldovich full bispectrum

Non-Gaussian (NG) statistics of the thermal Sunyaev-Zeldovich (tSZ) effect carry significant information which is not contained in the power spectrum. Here, we perform a joint Fisher analysis of the tSZ power spectrum and bispectrum to verify how much the full bispectrum can contribute to improve parameter constraints. We go beyond similar studies of this kind in several respects: first of all, we include the complete power spectrum and bispectrum (auto- and cross-) covariance in the analysis, computing all NG contributions; furthermore we consider a multi-component foreground scenario and model the effects of component separation in the forecasts; finally, we consider an extended set of both cosmological and intra-cluster medium parameters. We show that the tSZ bispectrum is very efficient at breaking parameter degeneracies, making it able to produce even stronger cosmological constraints than the tSZ power spectrum: e.g. the standard deviation on $\sigma_8$ shrinks from $\sigma^\text{PS}(\sigma_8)=0.35$ to $\sigma^\text{BS}(\sigma_8)=0.065$ when we consider a multi-parameter analysis. We find that this is mostly due to the different response of separate triangle types (e.g. equilateral and squeezed) to changes in model parameters. While weak, this shape dependence is clearly non-negligible for cosmological parameters, and it is even stronger, as expected, for intra-cluster medium parameters.

astro-ph.CO

Tomographic weak lensing bispectrum: a thorough analysis towards the next generation of galaxy surveys

We address key points for an efficient implementation of likelihood codes for modern weak lensing large-scale structure surveys. Specifically, we focus on the joint weak lensing convergence power spectrum-bispectrum probe and we tackle the numerical challenges required by a realistic analysis. Under the assumption of (multivariate) Gaussian likelihoods, we have developed a high performance code that allows highly parallelised prediction of the binned tomographic observables and of their joint non-Gaussian covariance matrix accounting for terms up to the 6-point correlation function and super-sample effects. This performance allows us to qualitatively address several interesting scientific questions. We find that the bispectrum provides an improvement in terms of signal-to-noise ratio (S/N) of about 10% on top of the power spectrum, making it a non-negligible source of information for future surveys. Furthermore, we are capable to test the impact of theoretical uncertainties in the halo model used to build our observables; with presently allowed variations we conclude that the impact is negligible on the S/N. Finally, we consider data compression possibilities to optimise future analyses of the weak lensing bispectrum. We find that, ignoring systematics, 5 equipopulated redshift bins are enough to recover the information content of a Euclid-like survey, with negligible improvement when increasing to 10 bins. We also explore principal component analysis and dependence on the triangle shapes as ways to reduce the numerical complexity of the problem.

astro-ph.CO

Testing Hu-Sawicki f(R) gravity with the Effective Field Theory approach

We show how to fully map a specific model of modified gravity into the Einstein-Boltzmann solver EFTCAMB. This approach consists in few steps and allows to obtain the cosmological phenomenology of a model with minimal effort. We discuss all these steps, from the solution of the dynamical equations for the cosmological background of the model to the use of the mapping relations to cast the model into the effective field theory language and use the latter to solve for perturbations. We choose the Hu-Sawicki f(R) model of gravity as our working example. After solving the background and performing the mapping, we interface the algorithm with EFTCAMB and take advantage of the effective field theory framework to integrate the full dynamics of linear perturbations, returning all quantities needed to accurately compare the model with observations. We discuss some observational signatures of this model, focusing on the linear growth of cosmic structures. In particular we present the behavior of $f\sigma_8$ and $E_G$ that, unlike the $\Lambda$CDM scenario, are generally scale dependent in addition to redshift dependent. Finally, we study the observational implications of the model by comparing its cosmological predictions to the Planck 2015 data, including CMB lensing, the WiggleZ galaxy survey and the CFHTLenS weak lensing survey measurements. We find that while WiggleZ data favor a non-vanishing value of the Hu-Sawicki model parameter, $\log_{10}(-f^0_{R})$, and consequently a large value of $\sigma_8$, CFHTLenS drags the estimate of $\log_{10}(-f^0_{R})$ back to the $\Lambda$CDM limit.

astro-ph.CO