arXiv ScienceSearch

arXiv subjects

Max E. Lee

Publications and source records attributed to Max E. Lee.

11 recordsLinked to original sources

BIND (Baryonic INpainting with Deep learning): A Field-level Emulator for Galaxy Groups and Clusters

Baryonic feedback is a dominant source of systematic uncertainty for upcoming weak-lensing surveys, but current tools for modeling its effect rely on spherical approximations and density profiles calibrated almost entirely on two-point statistics. We introduce BIND (Baryonic INpainting with Deep learning), a conditional flow-matching model that learns a field-level mapping from dark-matter-only halos to their hydrodynamical counterparts. BIND is trained on halos from the 1024 paired hydrodynamical and dark-matter-only simulations of the CAMELS $50\,h^{-1}\,\mathrm{Mpc}$ SB35 suite and samples dark matter, gas, and stellar mass fields over redshift across the full 35-dimensional $\Lambda$CDM and IllustrisTNG galaxy formation parameter space. BIND recovers dark matter, gas, and stellar masses at the percent level, reproduces azimuthally averaged profiles to $\lesssim10\%$ at all radii, and matches halo shape distributions with high fidelity. The learned parameter dependence captures the rank correlations between the generated fields and the subgrid parameters, and the field-level response to individual parameter variations is recovered in both sign and morphology. Halo mass is never supplied as conditioning, yet the baryon fraction, stellar-to-halo mass relation, inter-component scaling relations, and the joint covariance of their residuals are all reproduced. We finally show that, applied halo-by-halo to a $(50\,h^{-1}\,\mathrm{Mpc})^3$ $N$-body volume with $512^3$ particles, BIND reproduces the projected matter power spectrum suppression to the accuracy ceiling set by pasting in the hydrodynamical halos themselves, in minutes on one GPU. We release the trained BIND models and all generated halos as open-source tools. A companion paper extends BIND to thermodynamic fields and non-Gaussian weak-lensing statistics.

astro-ph.GA

BINDing the lightcone: A suite of astrophysical ray-traced weak lensing and SZ maps

Recent multiwavelength observations of galaxy group and cluster gas suggest stronger baryonic feedback than our best-calibrated hydrodynamical simulations produce, while modeling this feedback remains a primary source of uncertainty in Stage-IV weak-lensing (WL) analyses. We present a suite of ray-traced maps generated with BIND (Baryonic INpainting with Deep learning), a conditional flow-matching model that paints baryonic mass and gas thermodynamics onto the halos of dark-matter-only simulations, and which was developed in a companion paper. Applied to IllustrisTNG300-Dark and ray-traced, we generate convergence, optical depth, and Compton-$y$ maps at five source redshifts, each with $1000$ pseudo-independent realizations. We build lightcones across a 256-node Sobol sequence spanning the thirty-dimensional IllustrisTNG galaxy formation prior, with individual parameter variations and at the fiducial model. In validation, the maps match those built from the IllustrisTNG300 halos to within LSST-Y10-like precision for a range of WL statistics. Across the prior, the response to feedback exceeds Stage-IV statistical precision by more than an order of magnitude on small scales, and different statistics respond to different model sectors: galactic winds control the WL power spectrum and gas auto- and cross-spectra, while the stellar initial mass function slope and AGN parameters shape the morphological statistics (PDF, peaks, minima, and Minkowski functionals). Finally, we find that the response of statistics can be compressed into seven halo properties which linearly predict a range of WL and SZ statistics. We publicly release the maps, statistics, and model tables.

astro-ph.CO

Learning the Universe with the 2nd Generation of CAMELS: Varying 35 parameters of the IllustrisTNG model in (50Mpc/h)^3 boxes

We present a new set of 1,192 cosmological simulations as part of the CAMELS project, in which a space of 35 cosmological, astrophysical, and numerical parameters is explored around the fiducial IllustrisTNG model. The volume of each of these simulations is (50Mpc/h)^3, eight times larger than that of previous CAMELS simulations. This provides lower sample variance as well as access to more massive halos and more diverse environments. We focus this work on exploring the advantages these differences provide for parameter inference powered by neural networks. We generate training sets based on the matter power spectra, projected maps of the volumes, graphs representing galaxy spatial distributions, and thermodynamical properties of massive halos. We employ multilayer perceptrons, convolutional neural networks, graph neural networks, and Gaussian processes, respectively, to extract information on the simulation parameters from these inputs while comparing systematically to analogous results from our previous generation of (25Mpc/h)^3 simulations. We generally find that the new, larger volumes produce tighter marginal constraints on the parameters, to degrees that vary between the different inputs. The improvements, however, scale more weakly than with the square root of the increase in the amount of data (i.e., physical volume). We interpret this as originating either from information loss due to mode coupling or from complex degeneracies in parameter space. We also discuss the effects on statistics of the intergalactic medium temperature from four new parameters that are varied in these simulations, which control the amplitude and timing of the ionizing background radiation. We publicly release the simulation outputs and ancillary data at https://camels.readthedocs.io.

astro-ph.CO

Efficiently emulating distribution functions in gigaparsec volumes for varying cosmological parameters

We present a new method for emulating the halo mass function (HMF) and other distribution functions in large effective volumes, down to low halo masses, whilst simultaneously modifying large ranges of parameters, for a fraction of the cost of traditional periodic cosmological simulations. We demonstrate the method by selecting small regions, $V \sim (50 \,h^{-1}{\rm Mpc})^3$, with a range of overdensities from the Quijote suite, consisting of tens of thousands of $(1 \,h^{-1}{\rm Gpc})^3$ $N$-body simulation volumes run with varying $\Lambda$CDM parameters. We train a differentiable emulator, conditioned on the overdensity of the region and these global parameters, to reproduce the halo mass function in these regions. We then successfully recover the global distribution of halo masses of the entire box by integrating over the overdensity distribution. Our approach uses just $\sim\,$0.026% of the original simulation volume, and suggests that suites of targeted `zoom' simulations, extracted from low resolution parent volumes, can be used to emulate large volume simulations at a fraction of the computational cost, whilst simultaneously pushing the dynamic range to much lower masses than can be achieved in periodic simulations. We discuss emulation of other key dark matter and baryonic distribution functions, as well as higher order statistics, with implications for the interpretation of upcoming wide field surveys on observatories such as Euclid, Roman and Rubin.

astro-ph.CO

The impact of baryons on weak lensing statistics as a function of halo mass and radius

Upcoming weak lensing (WL) surveys such as those by {\it Euclid}, LSST, and {\it Roman} require percent-level control over systematic effects. A common approach to mitigating baryonic effects uses semi-analytic baryon correction models (BCMs) that modify halo profiles in dark matter-only (DMO) simulations, calibrated to statistics from hydrodynamic simulations. We investigate the limits of this approach by progressively replacing larger regions around halos of decreasing mass in DMO simulations with their hydrodynamical counterparts. We compare multiple statistics -- the matter ($P(k)$) and weak-lensing ($C_\ell$) power spectra, peak counts, minima, one-point PDFs, and Minkowski functionals -- from "Replace" fields against hydrodynamical and DMO simulations. We find that replacing all halos with $M\geq10^{12}\,h^{-1}\,{\rm M}_\odot$ out to $r\leq5R_{200}$ recovers $\sim 90\%$ of the baryonic suppression in $P(k)$ and $C_\ell$ with the remaining $\sim 10\%$ originating from lower-mass halos or material farther outside of DM halos. Each statistic has distinct sensitivities to baryons: $P(k)$ and $C_\ell$ are sensitive to a broad range of masses and radii, whereas WL peaks are primarily affected by the cores of massive halos. We show that BCMs applied to massive halos and calibrated to match hydrodynamical $P(k)$ make two cancelling "mistakes": they underpredict core masses and compensate by overpredicting baryonic impacts at larger radii, thereby explaining previously reported failures of peak statistics in these models. We provide a framework for diagnosing critical mass/radius regions in baryonic modeling for a range of statistics for next-generation BCMs.

astro-ph.CO

Cosmological back-reaction of baryons on dark matter in the CAMELS simulations

Baryonic processes such as radiative cooling and feedback from massive stars and active galactic nuclei (AGN) directly redistribute baryons in the Universe but also indirectly redistribute dark matter due to changes in the gravitational potential. In this work, we investigate this "back-reaction" of baryons on dark matter using thousands of cosmological hydrodynamic simulations from the Cosmology and Astrophysics with MachinE Learning Simulations (CAMELS) project, including parameter variations in the SIMBA, IllustrisTNG, ASTRID, and Swift-EAGLE galaxy formation models. Matching haloes to corresponding N-body (dark matter-only) simulations, we find that virial masses decrease owing to the ejection of baryons by feedback. Relative to N-body simulations, halo profiles show an increased dark matter density in the center (due to radiative cooling) and a decrease in density farther out (due to feedback), with both effects being strongest in SIMBA (> 450% increase at r < 0.01 Rvir). The clustering of dark matter strongly responds to changes in baryonic physics, with dark matter power spectra in some simulations from each model showing as much as 20% suppression or increase in power at k ~ 10 h/Mpc relative to N-body simulations. We find that the dark matter back-reaction depends intrinsically on cosmology (Omega_m and sigma_8) at fixed baryonic physics, and varies strongly with the details of the feedback implementation. These results emphasize the need for marginalizing over uncertainties in baryonic physics to extract cosmological information from weak lensing surveys as well as their potential to constrain feedback models in galaxy evolution.

astro-ph.CO

The effect of intrinsic alignments on weak lensing statistics in hydrodynamical simulations

The next generation of weak gravitational lensing surveys has the potential to place stringent constraints on cosmological parameters. However, their analysis is limited by systematics such as the intrinsic alignments of galaxies, which alter weak lensing convergence and can lead to biases in cosmological parameter estimations. In this work, we investigate the impact of intrinsic alignments on non-Gaussian statistics of the weak lensing field using galaxy shapes derived from the IllustrisTNG hydrodynamical simulation. We create two catalogs of ray-traced convergence maps: one that includes the measured intrinsic shape of each galaxy and another where all galaxies are randomly rotated to eliminate intrinsic alignments. We compare an exhaustive list of weak lensing statistics between the two catalogs, including the shear-shear correlation function, the map-level angular power spectrum, one-point, peak count, minimum distribution functions, and Minkowski functionals. For each statistic, we assess the level of statistical distinguishability between catalogs for a set of future survey angular areas. Our results reveal strong small-scale correlation in the alignment of galaxies and statistically significant boosts in weak lensing convergence in both positive and negative directions for high-significance peaks and minimums, respectively. Weak lensing analyses utilizing non-Gaussian statistics must account for intrinsic alignments to avoid significantly compromised cosmological inferences.

astro-ph.CO

Cosmological and Astrophysical Parameter Inference from Stacked Galaxy Cluster Profiles Using CAMELS-zoomGZ

We present a study on the inference of cosmological and astrophysical parameters using stacked galaxy cluster profiles. Utilizing the CAMELS-zoomGZ simulations, we explore how various cluster properties--such as X-ray surface brightness, gas density, temperature, metallicity, and Compton-y profiles--can be used to predict parameters within the 28-dimensional parameter space of the IllustrisTNG model. Through neural networks, we achieve a high correlation coefficient of 0.97 or above for all cosmological parameters, including $\Omega_{\rm m}$, $H_0$, and $\sigma_8$, and over 0.90 for the remaining astrophysical parameters, showcasing the effectiveness of these profiles for parameter inference. We investigate the impact of different radial cuts, with bins ranging from $0.1R_{200c}$ to $0.7R_{200c}$, to simulate current observational constraints. Additionally, we perform a noise sensitivity analysis, adding up to 40\% Gaussian noise (corresponding to signal-to-noise ratios as low as 2.5), revealing that key parameters such as $\Omega_{\rm m}$, $H_0$, and the IMF slope remain robust even under extreme noise conditions. We also compare the performance of full radial profiles against integrated quantities, finding that profiles generally lead to more accurate parameter inferences. Our results demonstrate that stacked galaxy cluster profiles contain crucial information on both astrophysical processes within groups and clusters and the underlying cosmology of the universe. This underscores their significance for interpreting the complex data expected from next-generation surveys and reveals, for the first time, their potential as a powerful tool for parameter inference.

astro-ph.CO

Zooming by in the CARPoolGP lane: new CAMELS-TNG simulations of zoomed-in massive halos

Galaxy formation models within cosmological hydrodynamical simulations contain numerous parameters with non-trivial influences over the resulting properties of simulated cosmic structures and galaxy populations. It is computationally challenging to sample these high dimensional parameter spaces with simulations, particularly for halos in the high-mass end of the mass function. In this work, we develop a novel sampling and reduced variance regression method, CARPoolGP, which leverages built-in correlations between samples in different locations of high dimensional parameter spaces to provide an efficient way to explore parameter space and generate low variance emulations of summary statistics. We use this method to extend the Cosmology and Astrophysics with MachinE Learning Simulations (CAMELS) to include a set of 768 zoom-in simulations of halos in the mass range of $10^{13} - 10^{14.5} M_\odot\,h^{-1}$ that span a 28-dimensional parameter space in the IllustrisTNG model. With these simulations and the CARPoolGP emulation method, we explore parameter trends in the Compton $Y-M$, black hole mass-halo mass, and metallicity-mass relations, as well as thermodynamic profiles and quenched fractions of satellite galaxies. We use these emulations to provide a physical picture of the complex interplay between supernova and active galactic nuclei feedback. We then use emulations of the $Y-M$ relation of massive halos to perform Fisher forecasts on astrophysical parameters for future Sunyaev-Zeldovich observations and find a significant improvement in forecasted constraints. We publicly release both the simulation suite and CARPoolGP software package.

astro-ph.GA

Comparing weak lensing peak counts in baryonic correction models to hydrodynamical simulations

Next-generation weak lensing (WL) surveys, such as by the Vera Rubin Observatory's LSST, the $\textit{Roman}$ Space Telescope, and the $\textit{Euclid}$ space mission, will supply vast amounts of data probing small, highly nonlinear scales. Extracting information from these scales requires higher-order statistics and the controlling of related systematics such as baryonic effects. To account for baryonic effects in cosmological analyses at reduced computational cost, semi-analytic baryonic correction models (BCMs) have been proposed. Here, we study the accuracy of BCMs for WL peak counts, a well studied, simple, and effective higher-order statistic. We compare WL peak counts generated from the full hydrodynamical simulation IllustrisTNG and a baryon-corrected version of the corresponding dark matter-only simulation IllustrisTNG-Dark. We apply galaxy shape noise expected at the depths reached by DES, KiDS, HSC, LSST, $\textit{Roman}$, and $\textit{Euclid}$. We find that peak counts in BCMs are (i) accurate at the percent level for peaks with $\mathrm{S/N}<4$, (ii) statistically indistinguishable from IllustrisTNG in most current and ongoing surveys, but (iii) insufficient for deep future surveys covering the largest solid angles, such as LSST and $\textit{Euclid}$. We find that BCMs match individual peaks accurately, but underpredict the amplitude of the highest peaks. We conclude that existing BCMs are a viable substitute for full hydrodynamical simulations in cosmological parameter estimation from beyond-Gaussian statistics for ongoing and future surveys with modest solid angles. For the largest surveys, BCMs need to be refined to provide a more accurate match, especially to the highest peaks.

astro-ph.CO

MADLens, a python package for fast and differentiable non-Gaussian lensing simulations

We present MADLens a python package for producing non-Gaussian lensing convergence maps at arbitrary source redshifts with unprecedented precision. MADLens is designed to achieve high accuracy while keeping computational costs as low as possible. A MADLens simulation with only $256^3$ particles produces convergence maps whose power agree with theoretical lensing power spectra up to $L{=}10000$ within the accuracy limits of HaloFit. This is made possible by a combination of a highly parallelizable particle-mesh algorithm, a sub-evolution scheme in the lensing projection, and a machine-learning inspired sharpening step. Further, MADLens is fully differentiable with respect to the initial conditions of the underlying particle-mesh simulations and a number of cosmological parameters. These properties allow MADLens to be used as a forward model in Bayesian inference algorithms that require optimization or derivative-aided sampling. Another use case for MADLens is the production of large, high resolution simulation sets as they are required for training novel deep-learning-based lensing analysis tools. We make the MADLens package publicly available under a Creative Commons License (https://github.com/VMBoehm/MADLens).

astro-ph.CO