arXiv ScienceSearch

arXiv subjects

Shy Genel

Publications and source records attributed to Shy Genel.

At least 19 recordsLinked to original sources

BIND (Baryonic INpainting with Deep learning): A Field-level Emulator for Galaxy Groups and Clusters

Baryonic feedback is a dominant source of systematic uncertainty for upcoming weak-lensing surveys, but current tools for modeling its effect rely on spherical approximations and density profiles calibrated almost entirely on two-point statistics. We introduce BIND (Baryonic INpainting with Deep learning), a conditional flow-matching model that learns a field-level mapping from dark-matter-only halos to their hydrodynamical counterparts. BIND is trained on halos from the 1024 paired hydrodynamical and dark-matter-only simulations of the CAMELS $50\,h^{-1}\,\mathrm{Mpc}$ SB35 suite and samples dark matter, gas, and stellar mass fields over redshift across the full 35-dimensional $\Lambda$CDM and IllustrisTNG galaxy formation parameter space. BIND recovers dark matter, gas, and stellar masses at the percent level, reproduces azimuthally averaged profiles to $\lesssim10\%$ at all radii, and matches halo shape distributions with high fidelity. The learned parameter dependence captures the rank correlations between the generated fields and the subgrid parameters, and the field-level response to individual parameter variations is recovered in both sign and morphology. Halo mass is never supplied as conditioning, yet the baryon fraction, stellar-to-halo mass relation, inter-component scaling relations, and the joint covariance of their residuals are all reproduced. We finally show that, applied halo-by-halo to a $(50\,h^{-1}\,\mathrm{Mpc})^3$ $N$-body volume with $512^3$ particles, BIND reproduces the projected matter power spectrum suppression to the accuracy ceiling set by pasting in the hydrodynamical halos themselves, in minutes on one GPU. We release the trained BIND models and all generated halos as open-source tools. A companion paper extends BIND to thermodynamic fields and non-Gaussian weak-lensing statistics.

astro-ph.GA

BINDing the lightcone: A suite of astrophysical ray-traced weak lensing and SZ maps

Recent multiwavelength observations of galaxy group and cluster gas suggest stronger baryonic feedback than our best-calibrated hydrodynamical simulations produce, while modeling this feedback remains a primary source of uncertainty in Stage-IV weak-lensing (WL) analyses. We present a suite of ray-traced maps generated with BIND (Baryonic INpainting with Deep learning), a conditional flow-matching model that paints baryonic mass and gas thermodynamics onto the halos of dark-matter-only simulations, and which was developed in a companion paper. Applied to IllustrisTNG300-Dark and ray-traced, we generate convergence, optical depth, and Compton-$y$ maps at five source redshifts, each with $1000$ pseudo-independent realizations. We build lightcones across a 256-node Sobol sequence spanning the thirty-dimensional IllustrisTNG galaxy formation prior, with individual parameter variations and at the fiducial model. In validation, the maps match those built from the IllustrisTNG300 halos to within LSST-Y10-like precision for a range of WL statistics. Across the prior, the response to feedback exceeds Stage-IV statistical precision by more than an order of magnitude on small scales, and different statistics respond to different model sectors: galactic winds control the WL power spectrum and gas auto- and cross-spectra, while the stellar initial mass function slope and AGN parameters shape the morphological statistics (PDF, peaks, minima, and Minkowski functionals). Finally, we find that the response of statistics can be compressed into seven halo properties which linearly predict a range of WL and SZ statistics. We publicly release the maps, statistics, and model tables.

astro-ph.CO

Inductive Biases in Field-Level Cosmological Inference from Galaxy Catalogs

We perform field-level likelihood-free inference of the matter density parameter $\Omega_m$ from simulated galaxy catalogs using machine learning models with differing inductive biases. Using hydrodynamic simulations from CAMELS, we examine how observable choice and architecture govern cosmological information extraction. We consider galaxy positions and line-of-sight peculiar velocities, separately and jointly, and compare permutation-invariant Deep Sets, implemented with either multilayer perceptrons (MLPs) or Kolmogorov-Arnold Networks (KANs), to graph neural networks (GNNs), which explicitly encode spatial relations. We test in-distribution and out-of-distribution (OOD) performance across simulations with different subgrid galaxy-formation prescriptions. Deep Sets infer $\Omega_m$ from velocities alone with mean relative errors of approximately $18\%$ in-distribution and $\sim25\%$ OOD, with KANs and MLPs achieving comparable performance. In contrast, the same set-based approach does not yield useful $\sigma_8$ predictions in either in-distribution or cross-suite tests. Adding positions does not improve Deep Sets, while GNNs infer $\Omega_m$ with mean relative errors of about $10\%$ in-distribution and $10$--$17\%$ OOD. These results indicate that peculiar velocities provide the dominant source of $\Omega_m$ information for set-based models in this setting, while spatial information is most effectively used by architectures that explicitly encode galaxy-galaxy relations. Because the velocity inputs are exact simulated peculiar velocities, applications to survey data will require validation under realistic velocity-measurement noise, selection effects, and survey geometry.

astro-ph.CO

Evaluating the flexibility of the MillenniumTNG galaxy formation model with multi-zoom re-simulations

In this study we introduce a new simulation campaign designed to understand how parameters that control star-formation and AGN feedback processes in cosmological hydrodynamical simulations impact observables such as the galaxy stellar-mass function (GSMF) and the gas fractions in large dark matter halos. These simulations are zoom-ins to halos selected from the MillenniumTNG (MTNG) simulation, and are run employing a novel multi-zoom approach which simultaneously re-simulates several sub-regions of a given large volume at a higher resolution than the background, thus reducing computational cost and imbalances in parallelization. We measure the GSMF and gas-fractions in halos for each of the re-simulations, and train Gaussian-process emulators on these quantities. The resulting emulators predict the GSMF and gas-fractions in halos with $\sim0.1\,\mathrm{dex}$ and $\sim 10\%$ precision respectively. Using the emulators we can simultaneously fit recent measurements of both quantities, in particular the lower gas fractions now observed even for comparatively massive clusters. Interestingly, we find a combination of parameters of the MTNG galaxy formation model that provides a qualitatively good fit to both the measured GSMF and gas fractions. This combination of parameters differs from the fiducial one mainly by requiring that stellar-feedback is significantly less energetic, and that kinetic AGN feedback events are significantly more energetic and rare. This finding implies that the MTNG model can be consistent with scenarios of strong feedback that remove large amounts of gas from groups and clusters, albeit we caution that we have not extensively examined the effect of these new parameters on many quantities for which MTNG made successful predictions.

astro-ph.GA

Learning the Universe with cosmological rescaling of merger trees and semi-analytic galaxy formation models

Learning cosmology from galaxy surveys requires large suites of simulations spanning the cosmological and astrophysical parameter space, yet hydrodynamical simulations of galaxy formation remain prohibitively expensive. Semi-analytic models offer an inexpensive, physically grounded alternative, but still require halo merger trees from $N$-body simulations, and densely sampling cosmological parameters in sufficient volume remains expensive. We address this by extending cosmological rescaling to operate directly on merger trees and applying it in the $\Omega_{\rm m}$-$\sigma_8$ plane, running the Santa Cruz semi-analytic model for galaxy formation on the rescaled trees to produce galaxy populations across new cosmological and astrophysical parameters at negligible additional cost. A novel halo-profile-based correction, controlled by a single free parameter, suppresses systematic bias in rescaled halo masses to below the per cent level. We apply the method to parameter estimation of $\Omega_{\rm m}$ and $\sigma_8$ given either the stellar mass function or the two-point correlation function, finding that as few as 64, and potentially fewer, base $N$-body simulations, rescaled to $\sim1000$ training samples, match the accuracy of 750 dedicated $N$-body simulations; rescaling to 3200 realisations improves the prediction of $\Omega_{\rm m}$ by $\sim25\%$. Rescaling all merger trees from a single CAMELS-SAM $N$-body simulation costs $\sim0.1$ CPUh, compared to several thousand CPUh to run the simulation itself. We demonstrate a practical route to obtaining predictions of galaxy summary statistics across cosmological and astrophysical parameters, even with a relatively small number of base $N$-body simulations.

astro-ph.CO

Learning the Universe with the 2nd Generation of CAMELS: Varying 35 parameters of the IllustrisTNG model in (50Mpc/h)^3 boxes

We present a new set of 1,192 cosmological simulations as part of the CAMELS project, in which a space of 35 cosmological, astrophysical, and numerical parameters is explored around the fiducial IllustrisTNG model. The volume of each of these simulations is (50Mpc/h)^3, eight times larger than that of previous CAMELS simulations. This provides lower sample variance as well as access to more massive halos and more diverse environments. We focus this work on exploring the advantages these differences provide for parameter inference powered by neural networks. We generate training sets based on the matter power spectra, projected maps of the volumes, graphs representing galaxy spatial distributions, and thermodynamical properties of massive halos. We employ multilayer perceptrons, convolutional neural networks, graph neural networks, and Gaussian processes, respectively, to extract information on the simulation parameters from these inputs while comparing systematically to analogous results from our previous generation of (25Mpc/h)^3 simulations. We generally find that the new, larger volumes produce tighter marginal constraints on the parameters, to degrees that vary between the different inputs. The improvements, however, scale more weakly than with the square root of the increase in the amount of data (i.e., physical volume). We interpret this as originating either from information loss due to mode coupling or from complex degeneracies in parameter space. We also discuss the effects on statistics of the intergalactic medium temperature from four new parameters that are varied in these simulations, which control the amplitude and timing of the ionizing background radiation. We publicly release the simulation outputs and ancillary data at https://camels.readthedocs.io.

astro-ph.CO

The Thermodynamic and Kinematic Evolution of Circumgalactic Gas around $z=1$ in the IllustrisTNG model

The circumgalactic medium (CGM) is known to contain multiphase gas in various stages of evolution and interaction with the galaxy. In order to characterize its detailed behavior on short timescales, we use a subregion of the TNG100 cosmological simulation to study the evolution of the $z=1$ CGM around six galaxies in $10^{11.5}-10^{12}$ $M_{\odot}$ halos at a high time cadence of $\approx2$ Myr. We use Monte Carlo tracer particles to follow this CGM gas forward in time in a Lagrangian way and determine how its thermodynamic and kinematic properties change. We find that CGM gas mixes between different temperature and density phases quickly and within $\approx500$ Myr evolves into distinct cold ($T\approx10^4$ $\rm{K}$) and warm-hot ($T\approx10^{5.5}$ $\rm{K}$) phases at small and large distances from the galaxy, respectively, regardless of its initial ($z=1$) halo-centric radius. This is largely driven by feedback from the galaxy, which heats and ejects cold gas that had previously cooled and accreted toward and occasionally into the galaxy from the outer CGM. We see signatures of this process in autocorrelations of kinematic quantities, which take $\approx400$ Myr to fully decorrelate from their initial values, suggesting a timescale over which feedback disrupts and reprocesses CGM gas. We also examine gas in narrow temperature and density ranges associated with commonly observed ions and find that gas that is O VI-like stays in its phase for hundreds of Myr longer than gas that is Mg II-like or C IV-like, suggesting that CGM observations of different species could probe gas in different evolutionary states, even if the gas is cospatial.

astro-ph.GA

Efficiently emulating distribution functions in gigaparsec volumes for varying cosmological parameters

We present a new method for emulating the halo mass function (HMF) and other distribution functions in large effective volumes, down to low halo masses, whilst simultaneously modifying large ranges of parameters, for a fraction of the cost of traditional periodic cosmological simulations. We demonstrate the method by selecting small regions, $V \sim (50 \,h^{-1}{\rm Mpc})^3$, with a range of overdensities from the Quijote suite, consisting of tens of thousands of $(1 \,h^{-1}{\rm Gpc})^3$ $N$-body simulation volumes run with varying $\Lambda$CDM parameters. We train a differentiable emulator, conditioned on the overdensity of the region and these global parameters, to reproduce the halo mass function in these regions. We then successfully recover the global distribution of halo masses of the entire box by integrating over the overdensity distribution. Our approach uses just $\sim\,$0.026% of the original simulation volume, and suggests that suites of targeted `zoom' simulations, extracted from low resolution parent volumes, can be used to emulate large volume simulations at a fraction of the computational cost, whilst simultaneously pushing the dynamic range to much lower masses than can be achieved in periodic simulations. We discuss emulation of other key dark matter and baryonic distribution functions, as well as higher order statistics, with implications for the interpretation of upcoming wide field surveys on observatories such as Euclid, Roman and Rubin.

astro-ph.CO

The impact of baryons on weak lensing statistics as a function of halo mass and radius

Upcoming weak lensing (WL) surveys such as those by {\it Euclid}, LSST, and {\it Roman} require percent-level control over systematic effects. A common approach to mitigating baryonic effects uses semi-analytic baryon correction models (BCMs) that modify halo profiles in dark matter-only (DMO) simulations, calibrated to statistics from hydrodynamic simulations. We investigate the limits of this approach by progressively replacing larger regions around halos of decreasing mass in DMO simulations with their hydrodynamical counterparts. We compare multiple statistics -- the matter ($P(k)$) and weak-lensing ($C_\ell$) power spectra, peak counts, minima, one-point PDFs, and Minkowski functionals -- from "Replace" fields against hydrodynamical and DMO simulations. We find that replacing all halos with $M\geq10^{12}\,h^{-1}\,{\rm M}_\odot$ out to $r\leq5R_{200}$ recovers $\sim 90\%$ of the baryonic suppression in $P(k)$ and $C_\ell$ with the remaining $\sim 10\%$ originating from lower-mass halos or material farther outside of DM halos. Each statistic has distinct sensitivities to baryons: $P(k)$ and $C_\ell$ are sensitive to a broad range of masses and radii, whereas WL peaks are primarily affected by the cores of massive halos. We show that BCMs applied to massive halos and calibrated to match hydrodynamical $P(k)$ make two cancelling "mistakes": they underpredict core masses and compensate by overpredicting baryonic impacts at larger radii, thereby explaining previously reported failures of peak statistics in these models. We provide a framework for diagnosing critical mass/radius regions in baryonic modeling for a range of statistics for next-generation BCMs.

astro-ph.CO

The Impact of Galaxy Formation on Galaxy Biasing, and Implications for Primordial non-Gaussianity Constraints

The parameter $f_{\textrm{NL}}$ measures the local non-Gaussianity in the primordial energy fluctuations of the Universe, with any deviation from $f_{\textrm{NL}}=0$ providing key constraints on inflationary models. Galaxy clustering is sensitive to $f_{\textrm{NL}}$ at large scale modes and the next generation of galaxy surveys will approach a statistical error of $\sigma_{f_{\textrm{NL}}}\sim1$. However, the systematic errors on these constraints are dominated by the degeneracy of $f_{\textrm{NL}}$ with the galaxy bias parameters $b_1$ (galaxy overdensities caused by mass perturbations) and $b_{\phi}$ (galaxy overdensities caused by primordial potential perturbations). It has been shown that the assumed scaling of $b_{\phi}(z)=2\delta_c (b_1(z)-1)$ is not accurate for realistically simulated galaxies, and depends both on the galaxy selection and the way that galaxies are modeled. To address this, we leverage the CAMELS-SAM pipeline to explore how varying parameters of galaxy formation affects $b_{\phi}$ and $b_1$ for various galaxy selections. We run separate-universe N-body simulations of $L=205 h^{-1}$ cMpc and $N=1280^3$ to measure $b_{\phi}$, and run 55 unique instances of the Santa Cruz semi-analytic model with varying parameters of stellar and AGN feedback. We find the behavior and evolution of a SC-SAM model's stellar-, SFR- and sSFR- to halo mass relationships track well with how $b_1$ and $b_{\phi}(b_1)$ change across redshift and selection for the SC-SAM. We find our variations of the SC-SAM encapsulate the $b_{\phi}$ behavior previously measured in IllustrisTNG, the Munich SAM, and Galacticus.Finally, we identify sSFR selections as particularly robust to varied galaxy modeling.

astro-ph.CO

Cosmological back-reaction of baryons on dark matter in the CAMELS simulations

Baryonic processes such as radiative cooling and feedback from massive stars and active galactic nuclei (AGN) directly redistribute baryons in the Universe but also indirectly redistribute dark matter due to changes in the gravitational potential. In this work, we investigate this "back-reaction" of baryons on dark matter using thousands of cosmological hydrodynamic simulations from the Cosmology and Astrophysics with MachinE Learning Simulations (CAMELS) project, including parameter variations in the SIMBA, IllustrisTNG, ASTRID, and Swift-EAGLE galaxy formation models. Matching haloes to corresponding N-body (dark matter-only) simulations, we find that virial masses decrease owing to the ejection of baryons by feedback. Relative to N-body simulations, halo profiles show an increased dark matter density in the center (due to radiative cooling) and a decrease in density farther out (due to feedback), with both effects being strongest in SIMBA (> 450% increase at r < 0.01 Rvir). The clustering of dark matter strongly responds to changes in baryonic physics, with dark matter power spectra in some simulations from each model showing as much as 20% suppression or increase in power at k ~ 10 h/Mpc relative to N-body simulations. We find that the dark matter back-reaction depends intrinsically on cosmology (Omega_m and sigma_8) at fixed baryonic physics, and varies strongly with the details of the feedback implementation. These results emphasize the need for marginalizing over uncertainties in baryonic physics to extract cosmological information from weak lensing surveys as well as their potential to constrain feedback models in galaxy evolution.

astro-ph.CO

The Cosmic Baryon Cycle in IllustrisTNG: flows of mass, energy, and metals

We measure and analyze the inflows and outflows of mass, energy, and metals through the interstellar medium (ISM) and circumgalactic medium (CGM) of galaxies in the IllustrisTNG100 simulations. We identify the dominant feedback mechanism in bins of halo virial mass and redshift by computing the integrated energy input from SNe and the ``kinetic'' and ``thermal'' mode of AGN feedback. We measure all quantities in a shell at the virial radius (``halo scale'') and one chosen to be approximately at the interface of the CGM and the interstellar medium (ISM; ``ISM scale''). We find that galaxies have strong net positive inflows on halo scales, and weaker but still net positive inflows on ISM scales, at $z\gtrsim 2$. At later times, partially due to the onset of kinetic AGN feedback in massive halos, inflows and outflows nearly balance one another, leading to the familiar effects of the slow-down of galaxy growth and the onset of quenching. Halos dominated by SN feedback show only weak evidence of preventative feedback on halo scales, and we see excess ISM scale accretion indicative of rapid gas recycling. Wind mass loadings decrease with increasing halo mass, and with increasing redshift, while energy loadings are nearly independent of both mass and redshift. The detailed catalogs of these mass, metal, and energy inflow and outflow rates on galaxy and halo scales can be used to guide empirical and semi-analytic models, and provide deeper insight into how galaxy growth and quenching is regulated in the IllustrisTNG simulations.

astro-ph.GA

On the Frequency of Multiple Galaxy Mergers in $\Lambda$CDM Cosmological Simulations

Mergers are believed to play a pivotal role in galaxy evolution, and measuring the galaxy merger fraction is a longstanding goal of both observational and theoretical studies. In this work, we extend the consideration of the merger fraction from the standard measure of binary mergers, namely those comprising two merging galaxies, to multiple mergers, namely mergers involving three or more galaxies. We use the Illustris and IllustrisTNG cosmological hydrodynamical simulations to provide a theoretical prediction for the fraction of galaxy systems that are involved in a multiple merger as a function of various parameters, with a focus on the relationship between the multiple merger fraction $f_m$ and the total merger fraction $f_t$. We generally find that binary mergers dominate the total fraction and that $f_m\approx (0.5-0.7)f_t^{5/3}$, a prediction that can be tested observationally. We further compare the empirical simulation results with toy models where mergers occur, on the evolution timeline of a galaxy, either at constant intervals or as a Poisson process at a constant rate. From these comparisons, where the toy models typically produce lower multiple merger fractions, we conclude that in cosmological simulations, mergers are more strongly clustered in time than in these toy scenarios, likely reflecting the hierarchical nature of cosmological structure formation.

astro-ph.GA

One latent to fit them all: a unified representation of baryonic feedback on matter distribution

Accurate and parsimonious quantification of baryonic feedback on matter distribution is of crucial importance for understanding both cosmology and galaxy formation from observational data. This is, however, challenging given the large discrepancy among different models of galaxy formation simulations, and their distinct subgrid physics parameterizations. Using 5,072 simulations from 4 different models covering broad ranges in their parameter spaces, we find a unified 2D latent representation. Compared to the simulations and other phenomenological models, our representation is independent of both time and cosmology, much lower-dimensional, and disentangled in its impacts on the matter power spectra. The common latent space facilitates the comparison of parameter spaces of different models and is readily interpretable by correlation with each. The two latent dimensions provide a complementary representation of baryonic effects, linking black hole and supernova feedback to distinct and interpretable impacts on both the matter power spectrum, and field, level. Our approach enables developing robust and economical analytic models for optimal gain of physical information from data, and is generalizable to other fields with significant modeling uncertainty.

astro-ph.CO

How does feedback affect the star formation histories of galaxies?

Star formation in galaxies is regulated by the interplay of a range of processes that shape the multiphase gas in the interstellar and circumgalactic media. Using the CAMELS suite of cosmological simulations, we study the effects of varying feedback and cosmology on the average star formation histories (SFHs) of galaxies at $z\sim0$ across the IllustrisTNG, SIMBA and ASTRID galaxy formation models. We find that galaxy SFHs in all three models are sensitive to changes in stellar feedback, which affects the efficiency of baryon cycling and the rates at which central black holes grow, while effects of varying AGN feedback depend on model-dependent implementations of black hole seeding, accretion and feedback. We also find strong interaction terms that couple stellar and AGN feedback, usually by regulating the amount of gas available for the central black hole to accrete. Using a double power-law to describe the average SFHs, we derive a general set of equations relating the shape of the SFHs to physical quantities like baryon fraction and black hole mass across all three models. We find that a single set of equations (albeit with different coefficients) can describe the SFHs across all three CAMELS models, with cosmology dominating the SFH at early times, followed by halo accretion, and feedback and baryon cycling at late times. Galaxy SFHs provide a novel, complementary probe to constrain cosmology and feedback, and can connect the observational constraints from current and upcoming galaxy surveys with the physical mechanisms responsible for regulating galaxy growth and quenching.

astro-ph.GA

Enhanced Star Formation and Black Hole Accretion Rates in Galaxy Mergers in IllustrisTNG50

Many theoretical and observational studies have suggested that galaxy mergers may trigger enhanced star formation or active galactic nuclei (AGN) activity. We present an analysis of merging and nonmerging galaxies from $0.2 \leq z \leq 3$ in the IllustrisTNG50 simulation. These galaxies encompass a range of masses ($M_\star > 10^{8}M_\odot$), multiple merger stages, and mass ratios ($\geq1:10$). We examine the effect that galaxy mergers have on star formation and black hole accretion rates in the TNG50 universe. We additionally investigate how galaxy and black hole mass, merger stage, merger mass ratio, and redshift affect these quantities. Mergers in our sample show excess specific star formation rates (sSFR) at $z \leq 3$ and enhanced specific black hole accretion rates (sBHAR) at $z \lesssim 2$. The difference between sSFRs and sBHARs in the merging sample compared to the non-merging sample increases as redshift decreases. Additionally, we show that these enhancements persist for at least $\sim1$ Gyr after the merger event. Investigating how mergers behave in the TNG50 simulation throughout cosmic time enables both a better appreciation of the importance of spatial resolution in cosmological simulations and a better basis to understand our high-$z$ universe with observations from $\textit{JWST}$.

astro-ph.GA

The effect of intrinsic alignments on weak lensing statistics in hydrodynamical simulations

The next generation of weak gravitational lensing surveys has the potential to place stringent constraints on cosmological parameters. However, their analysis is limited by systematics such as the intrinsic alignments of galaxies, which alter weak lensing convergence and can lead to biases in cosmological parameter estimations. In this work, we investigate the impact of intrinsic alignments on non-Gaussian statistics of the weak lensing field using galaxy shapes derived from the IllustrisTNG hydrodynamical simulation. We create two catalogs of ray-traced convergence maps: one that includes the measured intrinsic shape of each galaxy and another where all galaxies are randomly rotated to eliminate intrinsic alignments. We compare an exhaustive list of weak lensing statistics between the two catalogs, including the shear-shear correlation function, the map-level angular power spectrum, one-point, peak count, minimum distribution functions, and Minkowski functionals. For each statistic, we assess the level of statistical distinguishability between catalogs for a set of future survey angular areas. Our results reveal strong small-scale correlation in the alignment of galaxies and statistically significant boosts in weak lensing convergence in both positive and negative directions for high-significance peaks and minimums, respectively. Weak lensing analyses utilizing non-Gaussian statistics must account for intrinsic alignments to avoid significantly compromised cosmological inferences.

astro-ph.CO

Cosmology with One Galaxy: Auto-Encoding the Galaxy Properties Manifold

Cosmological simulations like CAMELS and IllustrisTNG characterize hundreds of thousands of galaxies using various internal properties. Previous studies have demonstrated that machine learning can be used to infer the cosmological parameter $\Omega_m$ from the internal properties of even a single randomly selected simulated galaxy. This ability was hypothesized to originate from galaxies occupying a low-dimensional manifold within a higher-dimensional galaxy property space, which shifts with variations in $\Omega_m$. In this work, we investigate how galaxies occupy the high-dimensional galaxy property space, particularly the effect of $\Omega_m$ and other cosmological and astrophysical parameters on the putative manifold. We achieve this by using an autoencoder with an Information-Ordered Bottleneck (IOB), a neural layer with adaptive compression, to perform dimensionality reduction on individual galaxy properties from CAMELS simulations, which are run with various combinations of cosmological and astrophysical parameters. We find that for an autoencoder trained on the fiducial set of parameters, the reconstruction error increases significantly when the test set deviates from fiducial values of $\Omega_m$ and $A_{\text{SN1}}$, indicating that these parameters shift galaxies off the fiducial manifold. In contrast, variations in other parameters such as $\sigma_8$ cause negligible error changes, suggesting galaxies shift along the manifold. These findings provide direct evidence that the ability to infer $\Omega_m$ from individual galaxies is tied to the way $\Omega_m$ shifts the manifold. Physically, this implies that parameters like $\sigma_8$ produce galaxy property changes resembling natural scatter, while parameters like $\Omega_m$ and $A_{\text{SN1}}$ create unsampled properties, extending beyond the natural scatter in the fiducial model.

astro-ph.CO