arXiv ScienceSearch

arXiv subjects

Peter Carr

Publications and source records attributed to Peter Carr.

At least 19 recordsLinked to original sources

When to Sell an Asset? - A Distribution Builder Approach

We consider the question of the optimal timing of the sale of an asset with stochastic dynamics. Our analysis is based on the method of the distribution builder introduced by Sharpe, Goldstein and Blythe [SGB00] for the purpose of optimal portfolio selection. Instead of specifying a utility function or risk aversion coefficient, this tool directly elicits the target distribution of the investor. We show how the problem of an optimal asset sale is in this setting linked to the problem of finding a Skorokhod embedding of a distribution into a diffusion process. In the case where the asset process follows a geometric Brownian motion and a specific family of distributions is targeted, one can observe a risk-return tradeoff.

q-fin.PR

Argoverse 2: Next Generation Datasets for Self-Driving Perception and Forecasting

We introduce Argoverse 2 (AV2) - a collection of three datasets for perception and forecasting research in the self-driving domain. The annotated Sensor Dataset contains 1,000 sequences of multimodal data, encompassing high-resolution imagery from seven ring cameras, and two stereo cameras in addition to lidar point clouds, and 6-DOF map-aligned pose. Sequences contain 3D cuboid annotations for 26 object categories, all of which are sufficiently-sampled to support training and evaluation of 3D perception models. The Lidar Dataset contains 20,000 sequences of unlabeled lidar point clouds and map-aligned pose. This dataset is the largest ever collection of lidar sensor data and supports self-supervised learning and the emerging task of point cloud forecasting. Finally, the Motion Forecasting Dataset contains 250,000 scenarios mined for interesting and challenging interactions between the autonomous vehicle and other actors in each local scene. Models are tasked with the prediction of future motion for "scored actors" in each scenario and are provided with track histories that capture object location, heading, velocity, and category. In all three datasets, each scenario contains its own HD Map with 3D lane and crosswalk geometry - sourced from data captured in six distinct cities. We believe these datasets will support new and existing machine learning research problems in ways that existing datasets do not. All datasets are released under the CC BY-NC-SA 4.0 license.

cs.CV

Robust Replication of Volatility and Hybrid Derivatives on Jump Diffusions

We price and replicate a variety of claims written on the log price $X$ and quadratic variation $[X]$ of a risky asset, modeled as a positive semimartingale, subject to stochastic volatility and jumps. The pricing and hedging formulas do not depend on the dynamics of volatility process, aside from integrability and independence assumptions; in particular, the volatility process may be non-Markovian and exhibit jumps of unknown distribution. The jump risk may be driven by any finite activity Poisson random measure with bounded jump sizes. As hedging instruments, we use the underlying risky asset, a zero-coupon bond, and European calls and puts with the same maturity as the claim to be hedged. Examples of contracts that we price include variance swaps, volatility swaps, a claim that pays the realized Sharpe ratio, and a call on a leveraged exchange traded fund.

q-fin.MF

Semi-closed form prices of barrier options in the time-dependent CEV and CIR models

We continue a series of papers where prices of the barrier options written on the underlying, which dynamics follows some one factor stochastic model with time-dependent coefficients and the barrier, are obtained in semi-closed form, see (Carr and Itkin, 2020, Itkin and Muravey, 2020). This paper extends this methodology to the CIR model for zero-coupon bonds, and to the CEV model for stocks which are used as the corresponding underlying for the barrier options. We describe two approaches. One is generalization of the method of heat potentials for the heat equation to the Bessel process, so we call it the method of Bessel potentials. We also propose a general scheme how to construct the potential method for any linear differential operator with time-independent coefficients. The second one is the method of generalized integral transform, which is also extended to the Bessel process. In all cases, a semi-closed solution means that first, we need to solve numerically a linear Volterra equation of the second kind, and then the option price is represented as a one-dimensional integral. We demonstrate that computationally our method is more efficient than both the backward and forward finite difference methods while providing better accuracy and stability. Also, it is shown that both method don't duplicate but rather compliment each other, as one provides very accurate results at small maturities, and the other one - at high maturities.

q-fin.CP

Semi-closed form solutions for barrier and American options written on a time-dependent Ornstein Uhlenbeck process

In this paper we develop a semi-closed form solutions for the barrier (perhaps, time-dependent) and American options written on the underlying stock which follows a time-dependent OU process with a log-normal drift. This model is equivalent to the familiar Hull-White model in FI, or a time dependent OU model in FX. Semi-closed form means that given the time-dependent interest rate, continuous dividend and volatility functions, one need to solve numerically a linear (for the barrier option) or nonlinear (for the American option) Fredholm equation of the first kind. After that the option prices in all cases are presented as one-dimensional integrals of combination of the above solutions and Jacobi theta functions. We also demonstrate that computationally our method is more efficient than the backward finite difference method used for solving these problems, and can also be as efficient as the forward finite difference solver while providing better accuracy and stability.

q-fin.PR

Argoverse: 3D Tracking and Forecasting with Rich Maps

We present Argoverse -- two datasets designed to support autonomous vehicle machine learning tasks such as 3D tracking and motion forecasting. Argoverse was collected by a fleet of autonomous vehicles in Pittsburgh and Miami. The Argoverse 3D Tracking dataset includes 360 degree images from 7 cameras with overlapping fields of view, 3D point clouds from long range LiDAR, 6-DOF pose, and 3D track annotations. Notably, it is the only modern AV dataset that provides forward-facing stereo imagery. The Argoverse Motion Forecasting dataset includes more than 300,000 5-second tracked scenarios with a particular vehicle identified for trajectory forecasting. Argoverse is the first autonomous vehicle dataset to include "HD maps" with 290 km of mapped lanes with geometric and semantic metadata. All data is released under a Creative Commons license at www.argoverse.org. In our baseline experiments, we illustrate how detailed map information such as lane direction, driveable area, and ground height improves the accuracy of 3D object tracking and motion forecasting. Our tracking and forecasting experiments represent only an initial exploration of the use of rich maps in robotic perception. We hope that Argoverse will enable the research community to explore these problems in greater depth.

cs.CV

Using Machine Learning to Predict Realized Variance

In this paper we formulate a regression problem to predict realized volatility by using option price data and enhance VIX-styled volatility indices' predictability and liquidity. We test algorithms including regularized regression and machine learning methods such as Feedforward Neural Networks (FNN) on S&P 500 Index and its option data. By conducting a time series validation we find that both Ridge regression and FNN can improve volatility indexing with higher prediction performance and fewer options required. The best approach found is to predict the difference between the realized volatility and the VIX-styled index's prediction rather than to predict the realized volatility directly, representing a successful combination of human learning and machine learning. We also discuss suitability of different regression algorithms for volatility indexing and applications of our findings.

q-fin.MF

A lognormal type stochastic volatility model with quadratic drift

This paper presents a novel one-factor stochastic volatility model where the instantaneous volatility of the asset log-return is a diffusion with a quadratic drift and a linear dispersion function. The instantaneous volatility mean reverts around a constant level, with a speed of mean reversion that is affine in the instantaneous volatility level. The steady-state distribution of the instantaneous volatility belongs to the class of Generalized Inverse Gaussian distributions. We show that the quadratic term in the drift is crucial to avoid moment explosions and to preserve the martingale property of the stock price process. Using a conveniently chosen change of measure, we relate the model to the class of polynomial diffusions. This remarkable relation allows us to develop a highly accurate option price approximation technique based on orthogonal polynomial expansions.

q-fin.MF

A model-free backward and forward nonlinear PDEs for implied volatility

We derive a backward and forward nonlinear PDEs that govern the implied volatility of a contingent claim whenever the latter is well-defined. This would include at least any contingent claim written on a positive stock price whose payoff at a possibly random time is convex. We also discuss suitable initial and boundary conditions for those PDEs. Finally, we demonstrate how to solve them numerically by using an iterative finite-difference approach.

q-fin.CP

ADOL - Markovian approximation of rough lognormal model

In this paper we apply Markovian approximation of the fractional Brownian motion (BM), known as the Dobric-Ojeda (DO) process, to the fractional stochastic volatility model where the instantaneous variance is modelled by a lognormal process with drift and fractional diffusion. Since the DO process is a semi-martingale, it can be represented as an \Ito diffusion. It turns out that in this framework the process for the spot price $S_t$ is a geometric BM with stochastic instantaneous volatility $\sigma_t$, the process for $\sigma_t$ is also a geometric BM with stochastic speed of mean reversion and time-dependent colatility of volatility, and the supplementary process $\calV_t$ is the Ornstein-Uhlenbeck process with time-dependent coefficients, and is also a function of the Hurst exponent. We also introduce an adjusted DO process which provides a uniformly good approximation of the fractional BM for all Hurst exponents $H \in [0,1]$ but requires a complex measure. Finally, the characteristic function (CF) of $\log S_t$ in our model can be found in closed form by using asymptotic expansion. Therefore, pricing options and variance swaps (by using a forward CF) can be done via FFT, which is much easier than in rough volatility models.

q-fin.MF

Geometric Local Variance Gamma model

This paper describes another extension of the Local Variance Gamma model originally proposed by P. Carr in 2008, and then further elaborated on by Carr and Nadtochiy, 2017 (CN2017), and Carr and Itkin, 2018 (CI2018). As compared with the latest version of the model developed in CI2018 and called the ELVG (the Expanded Local Variance Gamma model), here we provide two innovations. First, in all previous papers the model was constructed based on a Gamma time-changed {\it arithmetic} Brownian motion: with no drift in CI2017, and with drift in CI2018, and the local variance to be a function of the spot level only. In contrast, here we develop a {\it geometric} version of this model with drift. Second, in CN2017 the model was calibrated to option smiles assuming the local variance is a piecewise constant function of strike, while in CI2018 the local variance is a piecewise linear} function of strike. In this paper we consider 3 piecewise linear models: the local variance as a function of strike, the local variance as function of log-strike, and the local volatility as a function of strike (so, the local variance is a piecewise quadratic function of strike). We show that for all these new constructions it is still possible to derive an ordinary differential equation for the option price, which plays a role of Dupire's equation for the standard local volatility model, and, moreover, it can be solved in closed form. Finally, similar to CI2018, we show that given multiple smiles the whole local variance/volatility surface can be recovered which does not require solving any optimization problem. Instead, it can be done term-by-term by solving a system of non-linear algebraic equations for each maturity which is fast.

q-fin.PR

Generalizing Geometric Brownian Motion

To convert standard Brownian motion $Z$ into a positive process, Geometric Brownian motion (GBM) $e^{\beta Z_t}, \beta >0$ is widely used. We generalize this positive process by introducing an asymmetry parameter $ \alpha \geq 0$ which describes the instantaneous volatility whenever the process reaches a new low. For our new process, $\beta$ is the instantaneous volatility as prices become arbitrarily high. Our generalization preserves the positivity, constant proportional drift, and tractability of GBM, while expressing the instantaneous volatility as a randomly weighted $L^2$ mean of $\alpha$ and $\beta$. The running minimum and relative drawup of this process are also analytically tractable. Letting $\alpha = \beta$, our positive process reduces to Geometric Brownian motion. By adding a jump to default to the new process, we introduce a non-negative martingale with the same tractabilities. Assuming a security's dynamics are driven by these processes in risk neutral measure, we price several derivatives including vanilla, barrier and lookback options.

q-fin.MF

Domain Adaptation through Synthesis for Unsupervised Person Re-identification

Drastic variations in illumination across surveillance cameras make the person re-identification problem extremely challenging. Current large scale re-identification datasets have a significant number of training subjects, but lack diversity in lighting conditions. As a result, a trained model requires fine-tuning to become effective under an unseen illumination condition. To alleviate this problem, we introduce a new synthetic dataset that contains hundreds of illumination conditions. Specifically, we use 100 virtual humans illuminated with multiple HDR environment maps which accurately model realistic indoor and outdoor lighting. To achieve better accuracy in unseen illumination conditions we propose a novel domain adaptation technique that takes advantage of our synthetic data and performs fine-tuning in a completely unsupervised way. Our approach yields significantly higher accuracy than semi-supervised and unsupervised state-of-the-art methods, and is very competitive with supervised techniques.

cs.CV

Diversity Regularized Spatiotemporal Attention for Video-based Person Re-identification

Video-based person re-identification matches video clips of people across non-overlapping cameras. Most existing methods tackle this problem by encoding each video frame in its entirety and computing an aggregate representation across all frames. In practice, people are often partially occluded, which can corrupt the extracted features. Instead, we propose a new spatiotemporal attention model that automatically discovers a diverse set of distinctive body parts. This allows useful information to be extracted from all frames without succumbing to occlusions and misalignments. The network learns multiple spatial attention models and employs a diversity regularization term to ensure multiple models do not discover the same body part. Features extracted from local image regions are organized by spatial attention model and are combined using temporal attention. As a result, the network learns latent representations of the face, torso and other body parts using the best available image patches from the entire video sequence. Extensive evaluations on three datasets show that our framework outperforms the state-of-the-art approaches by large margins on multiple metrics.

cs.CV

An Expanded Local Variance Gamma model

The paper proposes an expanded version of the Local Variance Gamma model of Carr and Nadtochiy by adding drift to the governing underlying process. Still in this new model it is possible to derive an ordinary differential equation for the option price which plays a role of Dupire's equation for the standard local volatility model. It is shown how calibration of multiple smiles (the whole local volatility surface) can be done in such a case. Further, assuming the local variance to be a piecewise linear function of strike and piecewise constant function of time this ODE is solved in closed form in terms of Confluent hypergeometric functions. Calibration of the model to market smiles does not require solving any optimization problem and, in contrast, can be done term-by-term by solving a system of non-linear algebraic equations for each maturity, which is fast.

q-fin.CP

Pricing Variance Swaps on Time-Changed Markov Processes

We prove that the variance swap rate (fair strike) equals the price of a co-terminal European-style contract when the underlying is an exponential Markov process, time-changed by an arbitrary continuous stochastic clock, which has arbitrary correlation with the driving Markov process, provided that the payoff function $G$ of the European contract satisfies an ordinary integro-differential equation, which depends only on the dynamics of the Markov process, not on the clock. We present examples of Markov processes where the function $G$ that prices the variance swap can be computed explicitly. In general, the solutions $G$ are not contained in the logarithmic family previously obtained in the special case where the Markov process is a L\'evy process.

q-fin.MF

Coordinated Multi-Agent Imitation Learning

We study the problem of imitation learning from demonstrations of multiple coordinating agents. One key challenge in this setting is that learning a good model of coordination can be difficult, since coordination is often implicit in the demonstrations and must be inferred as a latent variable. We propose a joint approach that simultaneously learns a latent coordination model along with the individual policies. In particular, our method integrates unsupervised structure learning with conventional imitation learning. We illustrate the power of our approach on a difficult problem of learning multiple policies for fine-grained behavior modeling in team sports, where different players occupy different roles in the coordinated team strategy. We show that having a coordination model to infer the roles of players yields substantially improved imitation loss compared to conventional baselines.

cs.LG

Smooth Imitation Learning for Online Sequence Prediction

We study the problem of smooth imitation learning for online sequence prediction, where the goal is to train a policy that can smoothly imitate demonstrated behavior in a dynamic and continuous environment in response to online, sequential context input. Since the mapping from context to behavior is often complex, we take a learning reduction approach to reduce smooth imitation learning to a regression problem using complex function classes that are regularized to ensure smoothness. We present a learning meta-algorithm that achieves fast and stable convergence to a good policy. Our approach enjoys several attractive properties, including being fully deterministic, employing an adaptive learning rate that can provably yield larger policy improvements compared to previous approaches, and the ability to ensure stable convergence. Our empirical results demonstrate significant performance gains over previous approaches.

cs.LG