arXiv ScienceSearch

arXiv · 1909.11536

Extrapolated full waveform inversion with deep learning

Abstract

The lack of low frequency information and a good initial model can seriously affect the success of full waveform inversion (FWI), due to the inherent cycle skipping problem. Computational low frequency extrapolation is in principle the most direct way to address this issue. By considering bandwidth extension as a regression problem in machine learning, we propose an architecture of convolutional neural network (CNN) to automatically extrapolate the missing low frequencies without preprocessing and post-processing steps. The bandlimited recordings are the inputs of the CNN and, in our numerical experiments, a neural network trained from enough samples can predict a reasonable approximation to the seismograms in the unobserved low frequency band, both in phase and in amplitude. The numerical experiments considered are set up on simulated P-wave data. In extrapolated FWI (EFWI), the low-wavenumber components of the model are determined from the extrapolated low frequencies, before proceeding with a frequency sweep of the bandlimited data. The proposed deep-learning method of low-frequency extrapolation shows adequate generalizability for the initialization step of EFWI. Numerical examples show that the neural network trained on several submodels of the Marmousi model is able to predict the low frequencies for the BP 2004 benchmark model. Additionally, the neural network can robustly process seismic data with uncertainties due to the existence of noise, poorly-known source wavelet, and different finite-difference scheme in the forward modeling operator. Finally, this approach is not subject to the structural limitations of other methods for bandwidth extension, and seems to offer a tantalizing solution to the problem of properly initializing FWI.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Hongyu Sun, Laurent Demanet. 2019-12-13. Extrapolated full waveform inversion with deep learning. https://doi.org/10.1190/geo2019-0195.1

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Three dimensional non-singular mollified elastic dislocation theory for extended width fault zones and inhomogeneous boundary element models

Classical elastic dislocation theory (CEDT) has two challenges when applied to faulting problems: 1) fictitious on-fault stresses not defined by ordinary integration and 2) the geometric unreality of infinitely thin fault zones. We show that both can be resolved by a mollified elastic dislocation theory (MEDT) built on Cortez blob (Cortez, 2001) mollified displacement discontinuity Green's functions, which represent deformation across spatially distributed fault zones of finite scale epsilon and produce singularity-free displacements and stresses everywhere. Analytical integration of the mollified source solution over arbitrary planar triangular elements is done with AI, and the resulting closed-form solutions allow for the calculation of non-singular stresses across geometrically complex fault systems. Further, we demonstrate the decoupling of the fault-zone width scale epsilon from the mesh length scale h, the elasticity analog of a result established for regularized viscous flow Stokeslets (Ferranti and Cortez, 2024). We use these mollified kernels to demonstrate numerically stable collocation boundary element models including spatially extended fault zones, material property variations and non-planar topography.

physics.geo-ph

Identifying the approach of a major earthquake

By analyzing the seismicity in natural time and studying the evolution of the fluctuations of the entropy change of seismicity under time reversal for various scales of different length i (number of events), we can identify the approach of a major earthquake (EQ) occurrence. The current investigation is extended from 1984 until now for the seismicity in Japan.

physics.geo-ph

ADEPTS: An auto-differentiable framework for time-dependent nonlinear thermo-chemical mantle convection inversion

Time-dependent mantle-dynamics inversion must address the high dimensionality of the initial state, nonlinear rheology, and gradient propagation through long-term thermo-mechanical evolution. We develop ADEPTS, a two-dimensional staggered-grid finite-difference framework for mantle-dynamics inversion based on automatic differentiation. The forward model solves incompressible Stokes flow, temperature advection-diffusion, and compositional advection with temperature- and strain-rate-dependent viscosity and plastic yielding. For the nonlinear Stokes system, we compare two gradient strategies: unrolled differentiation through a fixed number of Picard iterations and implicit differentiation of the converged discrete residual equations. Numerical experiments show that unrolled differentiation remains stable even when the nonlinear solve is not fully converged, whereas implicit differentiation requires sufficiently accurate nonlinear solutions; otherwise gradient consistency and optimization convergence deteriorate. With sufficiently converged solves, implicit differentiation recovers accurate gradients and reconstruction quality comparable to unrolled differentiation. Joint thermo-chemical twin experiments show that ADEPTS can simultaneously recover a high-dimensional initial temperature field and low-dimensional physical parameters, including compositional density, reference viscosity, and stress exponent, while fitting final-time temperature, surface horizontal velocity, and surface normal stress. These results demonstrate the feasibility of differentiable time-dependent mantle-dynamics inversion and clarify the different convergence requirements of unrolled and implicit differentiation.

physics.geo-ph