arXiv ScienceSearch

arXiv subjects

Christian Reisswig

Publications and source records attributed to Christian Reisswig.

At least 19 recordsLinked to original sources

Elastic Token Compression for Pixel-Space Diffusion Transformers

Natural images concentrate their detail in a small fraction of the frame, yet diffusion models spend a full token on every patch, in every layer and at every timestep. The waste is largest in pixel-space models, with no autoencoder to absorb low-level redundancy first. Probing a pretrained pixel text-to-image transformer, we find its middle-block tokens redundant wherever the image is flat. The redundancy occupies connected, content-shaped regions, and exploiting it requires tokens with the same geometry. Cutting a Hilbert ordering of the patches provides them. Consecutive positions are always image neighbours, so any contiguous run is a connected region whose size and shape follow the content, and grouping in two dimensions becomes a cut in one. Existing reductions each lose part of this. Similarity merging scatters its groups, latent bottlenecks discard position, and skipping deletes what it should summarize. We cut where the model's features change most and pool each run into one region token. Our Region Token Interface (\method{}) adapts a diffusion model to these tokens, with the region count drawn at random during fine-tuning so one checkpoint serves every budget. \method{} leads prior reduction methods at matched budgets, matches dense quality at $2.0\times$ the speed, and stays close at $2.6\times$. The code and models are open-sourced at https://eduardzamfir.github.io/rti

cs.CV

BERTgrid: Contextualized Embedding for 2D Document Representation and Understanding

For understanding generic documents, information like font sizes, column layout, and generally the positioning of words may carry semantic information that is crucial for solving a downstream document intelligence task. Our novel BERTgrid, which is based on Chargrid by Katti et al. (2018), represents a document as a grid of contextualized word piece embedding vectors, thereby making its spatial structure and semantics accessible to the processing neural network. The contextualized embedding vectors are retrieved from a BERT language model. We use BERTgrid in combination with a fully convolutional network on a semantic instance segmentation task for extracting fields from invoices. We demonstrate its performance on tabulated line item and document header field extraction.

cs.CL

Chargrid-OCR: End-to-end Trainable Optical Character Recognition for Printed Documents using Instance Segmentation

We present an end-to-end trainable approach for Optical Character Recognition (OCR) on printed documents. Specifically, we propose a model that predicts a) a two-dimensional character grid (\emph{chargrid}) representation of a document image as a semantic segmentation task and b) character boxes for delineating character instances as an object detection task. For training the model, we build two large-scale datasets without resorting to any manual annotation - synthetic documents with clean labels and real documents with noisy labels. We demonstrate experimentally that our method, trained on the combination of these datasets, (i) outperforms previous state-of-the-art approaches in accuracy (ii) is easily parallelizable on GPU and is, therefore, significantly faster and (iii) is easy to train and adapt to a new domain.

cs.CV

Chargrid: Towards Understanding 2D Documents

We introduce a novel type of text representation that preserves the 2D layout of a document. This is achieved by encoding each document page as a two-dimensional grid of characters. Based on this representation, we present a generic document understanding pipeline for structured documents. This pipeline makes use of a fully convolutional encoder-decoder network that predicts a segmentation mask and bounding boxes. We demonstrate its capabilities on an information extraction task from invoices and show that it significantly outperforms approaches based on sequential text or document images.

cs.CL

Energetics and phasing of nonprecessing spinning coalescing black hole binaries

We present an improved numerical relativity (NR) calibration of the new effective-one-body (EOB) model for coalescing non precessing spinning black hole binaries recently introduced by Damour and Nagar [Physical Review D 90, 044018 (2014)]. We do so by comparing the EOB predictions to both the phasing and the energetics provided by two independent sets of NR data covering mass ratios $1\leq q \leq 9.989$ and dimensionless spin range $-0.95\leq \chi\leq +0.994$. One set of data is a subset of the Simulating eXtreme Spacetimes (SXS) catalog of public waveforms; the other set consists of new simulations obtained with the Llama code plus Cauchy Characteristic Evolution. We present the first systematic computation of the gauge-invariant relation between the binding energy and the total angular momentum, $E_{b}(j)$, for a large sample of, spin-aligned, SXS and Llama data. The dynamics of the EOB model presented here involves only two free functional parameters, one ($a_6^c(\nu)$) entering the non spinning sector, as a 5PN effective correction to the interaction potential, and one ($c_3(\tilde{a}_1,\tilde{a}_2,\nu))$ in the spinning sector, as an effective next-to-next-to-next-to-leading order correction to the spin-orbit coupling. These parameters are determined (together with a third functional parameter $\Delta t_{\rm NQC}(\chi)$ entering the waveform) by comparing the EOB phasing with the SXS phasing, the consistency of the energetics being checked afterwards. The quality of the analytical model for gravitational wave data analysis purposes is assessed by computing the EOB/NR faithfulness. Over the NR data sample and when varying the total mass between 20 and 200~$M_\odot$ the EOB/NR unfaithfulness (integrated over the NR frequency range) is found to vary between $99.493\%$ and $99.984\%$ with a median value of $99.944\%$.

gr-qc

Comparing Gravitational Waveform Extrapolation to Cauchy-Characteristic Extraction in Binary Black Hole Simulations

We extract gravitational waveforms from numerical simulations of black hole binaries computed using the Spectral Einstein Code. We compare two extraction methods: direct construction of the Newman-Penrose (NP) scalar $\Psi_4$ at a finite distance from the source and Cauchy-characteristic extraction (CCE). The direct NP approach is simpler than CCE, but NP waveforms can be contaminated by near-zone effects---unless the waves are extracted at several distances from the source and extrapolated to infinity. Even then, the resulting waveforms can in principle be contaminated by gauge effects. In contrast, CCE directly provides, by construction, gauge-invariant waveforms at future null infinity. We verify the gauge invariance of CCE by running the same physical simulation using two different gauge conditions. We find that these two gauge conditions produce the same CCE waveforms but show differences in extrapolated-$\Psi_4$ waveforms. We examine data from several different binary configurations and measure the dominant sources of error in the extrapolated-$\Psi_4$ and CCE waveforms. In some cases, we find that NP waveforms extrapolated to infinity agree with the corresponding CCE waveforms to within the estimated error bars. However, we find that in other cases extrapolated and CCE waveforms disagree, most notably for $m=0$ "memory" modes.

gr-qc

The gravitational wave strain in the characteristic formalism of numerical relativity

The extraction of the gravitational wave signal, within the context of a characteristic numerical evolution is revisited. A formula for the gravitational wave strain is developed and tested, and is made publicly available as part of the PITT code within the Einstein Toolkit. Using the new strain formula, we show that artificial non-linear drifts inherent in time integrated waveforms can be reduced for the case of a binary black hole merger configuration. For the test case of a rapidly spinning stellar core collapse model, however, we find that the drift must have different roots.

gr-qc

Error-analysis and comparison to analytical models of numerical waveforms produced by the NRAR Collaboration

The Numerical-Relativity-Analytical-Relativity (NRAR) collaboration is a joint effort between members of the numerical relativity, analytical relativity and gravitational-wave data analysis communities. The goal of the NRAR collaboration is to produce numerical-relativity simulations of compact binaries and use them to develop accurate analytical templates for the LIGO/Virgo Collaboration to use in detecting gravitational-wave signals and extracting astrophysical information from them. We describe the results of the first stage of the NRAR project, which focused on producing an initial set of numerical waveforms from binary black holes with moderate mass ratios and spins, as well as one non-spinning binary configuration which has a mass ratio of 10. All of the numerical waveforms are analysed in a uniform and consistent manner, with numerical errors evaluated using an analysis code created by members of the NRAR collaboration. We compare previously-calibrated, non-precessing analytical waveforms, notably the effective-one-body (EOB) and phenomenological template families, to the newly-produced numerical waveforms. We find that when the binary's total mass is ~100-200 solar masses, current EOB and phenomenological models of spinning, non-precessing binary waveforms have overlaps above 99% (for advanced LIGO) with all of the non-precessing-binary numerical waveforms with mass ratios <= 4, when maximizing over binary parameters. This implies that the loss of event rate due to modelling error is below 3%. Moreover, the non-spinning EOB waveforms previously calibrated to five non-spinning waveforms with mass ratio smaller than 6 have overlaps above 99.7% with the numerical waveform with a mass ratio of 10, without even maximizing on the binary parameters.

gr-qc

The Transient Gravitational-Wave Sky

Interferometric detectors will very soon give us an unprecedented view of the gravitational-wave sky, and in particular of the explosive and transient Universe. Now is the time to challenge our theoretical understanding of short-duration gravitational-wave signatures from cataclysmic events, their connection to more traditional electromagnetic and particle astrophysics, and the data analysis techniques that will make the observations a reality. This paper summarizes the state of the art, future science opportunities, and current challenges in understanding gravitational-wave transients.

gr-qc

GRHydro: A new open source general-relativistic magnetohydrodynamics code for the Einstein Toolkit

We present the new general-relativistic magnetohydrodynamics (GRMHD) capabilities of the Einstein Toolkit, an open-source community-driven numerical relativity and computational relativistic astrophysics code. The GRMHD extension of the Toolkit builds upon previous releases and implements the evolution of relativistic magnetised fluids in the ideal MHD limit in fully dynamical spacetimes using the same shock-capturing techniques previously applied to hydrodynamical evolution. In order to maintain the divergence-free character of the magnetic field, the code implements both hyperbolic divergence cleaning and constrained transport schemes. We present test results for a number of MHD tests in Minkowski and curved spacetimes. Minkowski tests include aligned and oblique planar shocks, cylindrical explosions, magnetic rotors, Alfv\'en waves and advected loops, as well as a set of tests designed to study the response of the divergence cleaning scheme to numerically generated monopoles. We study the code's performance in curved spacetimes with spherical accretion onto a black hole on a fixed background spacetime and in fully dynamical spacetimes by evolutions of a magnetised polytropic neutron star and of the collapse of a magnetised stellar core. Our results agree well with exact solutions where these are available and we demonstrate convergence. All code and input files used to generate the results are available on http://einsteintoolkit.org. This makes our work fully reproducible and provides new users with an introduction to applications of the code.

gr-qc

General relativistic null-cone evolutions with a high-order scheme

We present a high-order scheme for solving the full non-linear Einstein equations on characteristic null hypersurfaces using the framework established by Bondi and Sachs. This formalism allows asymptotically flat spaces to be represented on a finite, compactified grid, and is thus ideal for far-field studies of gravitational radiation. We have designed an algorithm based on 4th-order radial integration and finite differencing, and a spectral representation of angular components. The scheme can offer significantly more accuracy with relatively low computational cost compared to previous methods as a result of the higher-order discretization. Based on a newly implemented code, we show that the new numerical scheme remains stable and is convergent at the expected order of accuracy.

gr-qc

The NINJA-2 catalog of hybrid post-Newtonian/numerical-relativity waveforms for non-precessing black-hole binaries

The Numerical INJection Analysis (NINJA) project is a collaborative effort between members of the numerical relativity and gravitational wave data analysis communities. The purpose of NINJA is to study the sensitivity of existing gravitational-wave search and parameter-estimation algorithms using numerically generated waveforms, and to foster closer collaboration between the numerical relativity and data analysis communities. The first NINJA project used only a small number of injections of short numerical-relativity waveforms, which limited its ability to draw quantitative conclusions. The goal of the NINJA-2 project is to overcome these limitations with long post-Newtonian - numerical relativity hybrid waveforms, large numbers of injections, and the use of real detector data. We report on the submission requirements for the NINJA-2 project and the construction of the waveform catalog. Eight numerical relativity groups have contributed 63 hybrid waveforms consisting of a numerical portion modelling the late inspiral, merger, and ringdown stitched to a post-Newtonian portion modelling the early inspiral. We summarize the techniques used by each group in constructing their submissions. We also report on the procedures used to validate these submissions, including examination in the time and frequency domains and comparisons of waveforms from different groups against each other. These procedures have so far considered only the $(\ell,m)=(2,2)$ mode. Based on these studies we judge that the hybrid waveforms are suitable for NINJA-2 studies. We note some of the plans for these investigations.

gr-qc

Energy versus Angular Momentum in Black Hole Binaries

Using accurate numerical relativity simulations of (nonspinning) black-hole binaries with mass ratios 1:1, 2:1 and 3:1 we compute the gauge invariant relation between the (reduced) binding energy $E$ and the (reduced) angular momentum $j$ of the system. We show that the relation $E(j)$ is an accurate diagnostic of the dynamics of a black-hole binary in a highly relativistic regime. By comparing the numerical-relativity $E^{\rm NR} (j)$ curve with the predictions of several analytic approximation schemes, we find that, while the usual, non-resummed post-Newtonian-expanded $E^{\rm PN} (j)$ relation exhibits large and growing deviations from $E^{\rm NR} (j)$, the prediction of the effective one-body formalism, based purely on known analytical results (without any calibration to numerical relativity), agrees strikingly well with the numerical-relativity results.

gr-qc

Initial data transients in binary black hole evolutions

We describe a method for initializing characteristic evolutions of the Einstein equations using a linearized solution corresponding to purely outgoing radiation. This allows for a more consistent application of the characteristic (null cone) techniques for invariantly determining the gravitational radiation content of numerical simulations. In addition, we are able to identify the {\em ingoing} radiation contained in the characteristic initial data, as well as in the initial data of the 3+1 simulation. We find that each component leads to a small but long lasting (several hundred mass scales) transient in the measured outgoing gravitational waves.

gr-qc

Gravitational Wave Extraction in Simulations of Rotating Stellar Core Collapse

We perform simulations of general relativistic rotating stellar core collapse and compute the gravitational waves (GWs) emitted in the core bounce phase of three representative models via multiple techniques. The simplest technique, the quadrupole formula (QF), estimates the GW content in the spacetime from the mass quadrupole tensor. It is strictly valid only in the weak-field and slow-motion approximation. For the first time, we apply GW extraction methods in core collapse that are fully curvature-based and valid for strongly radiating and highly relativistic sources. We employ three extraction methods computing (i) the Newman-Penrose (NP) scalar Psi_4, (ii) Regge-Wheeler-Zerilli-Moncrief (RWZM) master functions, and (iii) Cauchy-Characteristic Extraction (CCE) allowing for the extraction of GWs at future null infinity, where the spacetime is asymptotically flat and the GW content is unambiguously defined. The latter technique is the only one not suffering from residual gauge and finite-radius effects. All curvature-based methods suffer from strong non-linear drifts. We employ the fixed-frequency integration technique as a high-pass waveform filter. Using the CCE results as a benchmark, we find that finite-radius NP extraction yields results that agree nearly perfectly in phase, but differ in amplitude by ~1-7% at core bounce, depending on the model. RWZM waveforms, while in general agreeing in phase, contain spurious high-frequency noise of comparable amplitudes to those of the relatively weak GWs emitted in core collapse. We also find remarkably good agreement of the waveforms obtained from the QF with those obtained from CCE. They agree very well in phase but systematically underpredict peak amplitudes by ~5-11% which is comparable to the NP results and is within the uncertainties associated with core collapse physics. (abridged)

gr-qc

Notes on the integration of numerical relativity waveforms

A primary goal of numerical relativity is to provide estimates of the wave strain, $h$, from strong gravitational wave sources, to be used in detector templates. The simulations, however, typically measure waves in terms of the Weyl curvature component, $\psi_4$. Assuming Bondi gauge, transforming to the strain $h$ reduces to integration of $\psi_4$ twice in time. Integrations performed in either the time or frequency domain, however, lead to secular non-linear drifts in the resulting strain $h$. These non-linear drifts are not explained by the two unknown integration constants which can at most result in linear drifts. We identify a number of fundamental difficulties which can arise from integrating finite length, discretely sampled and noisy data streams. These issues are an artifact of post-processing data. They are independent of the characteristics of the original simulation, such as gauge or numerical method used. We suggest, however, a simple procedure for integrating numerical waveforms in the frequency domain, which is effective at strongly reducing spurious secular non-linear drifts in the resulting strain.

gr-qc

Gravitational memory in binary black hole mergers

In addition to the dominant oscillatory gravitational wave signals produced during binary inspirals, a non-oscillatory component arises from the nonlinear "memory" effect, sourced by the emitted gravitational radiation. The memory grows significantly during the late inspiral and merger, modifying the signal by an almost step-function profile, and making it difficult to model by approximate methods. We use numerical evolutions of binary black holes to evaluate the nonlinear memory during late-inspiral, merger and ringdown. We identify two main components of the signal: the monotonically growing portion corresponding to the memory, and an oscillatory part which sets in roughly at the time of merger and is due to the black hole ringdown. Counter-intuitively, the ringdown is most prominent for models with the lowest total spin. Thus, the case of maximally spinning black holes anti-aligned to the orbital angular momentum exhibits the highest signal-to-noise (SNR) for interferometric detectors. The largest memory offset, however, occurs for highly spinning black holes, with an estimated value of h^tot_20 \approx 0.24 in the maximally spinning case. These results are central to determining the detectability of nonlinear memory through pulsar timing array measurements.

gr-qc

High accuracy binary black hole simulations with an extended wave zone

We present results from a new code for binary black hole evolutions using the moving-puncture approach, implementing finite differences in generalised coordinates, and allowing the spacetime to be covered with multiple communicating non-singular coordinate patches. Here we consider a regular Cartesian near zone, with adapted spherical grids covering the wave zone. The efficiencies resulting from the use of adapted coordinates allow us to maintain sufficient grid resolution to an artificial outer boundary location which is causally disconnected from the measurement. For the well-studied test-case of the inspiral of an equal-mass non-spinning binary (evolved for more than 8 orbits before merger), we determine the phase and amplitude to numerical accuracies better than 0.010% and 0.090% during inspiral, respectively, and 0.003% and 0.153% during merger. The waveforms, including the resolved higher harmonics, are convergent and can be consistently extrapolated to $r\to\infty$ throughout the simulation, including the merger and ringdown. Ringdown frequencies for these modes (to $(\ell,m)=(6,6)$) match perturbative calculations to within 0.01%, providing a strong confirmation that the remnant settles to a Kerr black hole with irreducible mass $M_{\rm irr} = 0.884355\pm20\times10^{-6}$ and spin $S_f/M_f^2 = 0.686923 \pm 10\times10^{-6}$

gr-qc