arXiv Science⌕ Search

arXiv subjects

Lokman A Abbas-Turki

Publications and source records attributed to Lokman A Abbas-Turki.

2 recordsLinked to original sources

Stochastic Policy Gradient Methods in the Uncertain Volatility Model

The multidimensional Uncertain Volatility Model leads to robust option pricing problems under joint volatility and correlation uncertainty. Their numerical resolution quickly becomes challenging because the associated stochastic control problem is high-dimensional. We propose a backward actor-critic stochastic policy-gradient scheme tailored to this setting. The method combines a discrete dynamic programming principle with Proximal Policy Optimization and one-hidden-layer neural-network approximations of both the value function and the control policy. A key ingredient is the policy parameterization: continuous controls are represented through a squashed Gaussian policy built on a $C$-vine representation of correlation matrices, which enforces positive definiteness by construction. Beyond the robust price itself, the spatial gradient of the trained critic provides an approximation of the associated superhedging strategy. We assess the quality of this learned gradient through a dual formulation, which also yields a numerical dual estimate of the price. Numerical experiments on a range of multidimensional derivatives show that the method yields accurate prices, remains computationally efficient, and compares favorably with existing Monte Carlo and machine-learning-based benchmarks for robust pricing in the Uncertain Volatility Model.

q-fin.CP↗

A posteriori error bounds for the Uncertain Volatility Model

We develop a posteriori primal-dual bounds for numerical approximations of European option prices in the Uncertain Volatility Model. Given a smooth candidate approximation of the value function, a feedback control induced by its Hessian yields a primal lower bound. On the dual side, we derive a representation based on a matrix-valued Gamma field, which provides an upper bound through a nonnegative Hamiltonian penalty. We further show that, when the dual field is generated by the candidate itself, this upper bound admits an equivalent representation in terms of the residual of the associated Black-Scholes-Barenblatt equation. We then study discrete-time approximations of these primal and dual quantities and quantify the corresponding discretization errors. Finally, we investigate their numerical evaluation for candidates obtained by stochastic policy-gradient and physics-informed neural-network methods, and compare the resulting estimates with a martingale dual approach from the stochastic-control literature. The numerical experiments highlight the importance of derivative accuracy, and in particular of second-order information, for obtaining tight a posteriori bounds.

math.OC↗