arXiv ScienceSearch

arXiv · 1104.2700

N-body simulation for self-gravitating collisional systems with a new SIMD instruction set extension to the x86 architecture, Advanced Vector eXtensions

Abstract

We present a high-performance N-body code for self-gravitating collisional systems accelerated with the aid of a new SIMD instruction set extension of the x86 architecture: Advanced Vector eXtensions (AVX), an enhanced version of the Streaming SIMD Extensions (SSE). With one processor core of Intel Core i7-2600 processor (8 MB cache and 3.40 GHz) based on Sandy Bridge micro-architecture, we implemented a fourth-order Hermite scheme with individual timestep scheme (Makino and Aarseth, 1992), and achieved the performance of 20 giga floating point number operations per second (GFLOPS) for double-precision accuracy, which is two times and five times higher than that of the previously developed code implemented with the SSE instructions (Nitadori et al., 2006b), and that of a code implemented without any explicit use of SIMD instructions with the same processor core, respectively. We have parallelized the code by using so-called NINJA scheme (Nitadori et al., 2006a), and achieved 90 GFLOPS for a system containing more than N = 8192 particles with 8 MPI processes on four cores. We expect to achieve about 10 tera FLOPS (TFLOPS) for a self-gravitating collisional system with N 105 on massively parallel systems with at most 800 cores with Sandy Bridge micro-architecture. This performance will be comparable to that of Graphic Processing Unit (GPU) cluster systems, such as the one with about 200 Tesla C1070 GPUs (Spurzem et al., 2010). This paper offers an alternative to collisional N-body simulations with GRAPEs and GPUs.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Ataru Tanikawa, Kohji Yoshikawa, Takashi Okamoto, Keigo Nitadori. 2011-09-06. N-body simulation for self-gravitating collisional systems with a new SIMD instruction set extension to the x86 architecture, Advanced Vector eXtensions. https://doi.org/10.1016/j.newast.2011.07.001

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Observations of Binary Stars with the 1.3-m Devasthal Fast Optical Telescope Using Speckle Interferometry: An Attempt

We present a feasibility study of implementing optical interferometry and speckle techniques with the 1.3-m Devasthal Fast Optical Telescope (DFOT) at Aryabhatta Research Institute of Observational Sciences (ARIES), which is currently dedicated to photometric observations. Using the sCMOS camera at the DFOT backend, we perform interferometric speckle observations of six binary stars. Standard Speckle Interferometry algorithms are applied to analyze the recorded speckle images. While this study does not aim to achieve the diffraction limit of DFOT or address a full science-driven resolution case, it serves as a crucial testbed for instrumentation, data acquisition, and analysis of Speckles with DFOT. Although the individual speckle patterns exhibit significant positional variations from frame to frame, the Speckle Interferometry analysis successfully recovers the characteristic three-peak structure of the wide binaries, allowing their relative separations and position angles to be determined, demonstrating the viability of the approach. The measured angular relative separations of $γ$ Leo, $γ$ Vir, and $ζ$ Her are $4.744'' \pm 0.012''$, $3.387'' \pm 0.027''$, and $1.582'' \pm 0.026''$, respectively, with corresponding projected position angles of $126.94^{\circ} \pm 0.14^\circ$, $171.57^{\circ} \pm 0.33^\circ$, and $81.06^{\circ} \pm 0.91^\circ$. These measurements are consistent with previously reported values in the literature, which provides strong motivation for more systematic observations and future implementation of optical interferometry techniques at meter-class telescopes.

astro-ph.IM

Time-Domain Synthesis of Gravitational-Wave Detector Glitches using Class-Conditional Derivative Generative Adversarial Networks

Gravitational-wave detectors such as LIGO, Virgo, and KAGRA are highly sensitive instruments susceptible to many noise sources. Short-duration transient noise events, known as glitches, pose a particular challenge for data analysis pipelines, as they can mimic or obscure astrophysical signals. We present GlitchGAN, a class-conditional generative model built on the Conditional Derivative GAN (cDVGAN) architecture, capable of synthesizing seven glitch types from LIGO's third observing run (O3) directly in the time domain. GlitchGAN generalizes effectively, learning to reproduce a diverse glitch space consistent with high-quality DeepExtractor reconstructions, and can generate hybrid glitch morphologies by interpolating across its class-conditioning vector. It generates 1000 glitches in under 22 seconds on a CPU, suitable for detector simulations, mock data challenges, and pipeline validation. Synthetic glitches are validated using the Gravity Spy classifier and UMAP embeddings, both showing strong agreement with real data. To probe residual distributional differences, we train a separate holdout GlitchGAN model and use a downstream CNN to distinguish held-out real glitches from synthetic ones: detectability is high in a clean representation but drops substantially once both populations are injected into realistic detector noise, the condition under which they would typically be used. Despite this, GlitchGAN-generated glitches remain practically useful: augmenting real training sets with synthetic samples matches simple duplication of real data when data is abundant, and increasingly outperforms it as real data becomes scarce. Finally, we highlight a limitation of magnitude-only spectrograms: magnitude Q-transform classifiers can confidently misclassify physically unrealistic glitches from less robust models, underscoring the need for validation methods that preserve phase information.

astro-ph.IM

SGN: A python framework for stream-processing pipelines

We present the Stream Graph Navigator (SGN), a lightweight Python framework for building streaming data applications. In SGN, stream-processing pipelines are built by connecting computational components into directed acyclic graphs that run within an event loop. The time-series extension of the SGN library, SGN-TS, introduces signal processing methods to handle time series data. Together, SGN and SGN-TS provide the foundation for SGNL, a matched-filtering gravitational-wave search pipeline, and are being adopted by multiple projects across the low-latency gravitational-wave data analysis infrastructure as an extensible and maintainable framework for future gravitational-wave observations.

astro-ph.IM