We quantify the cosmological spread of baryons relative to their initial neighbouring dark matter distribution using thousands of state-of-the-art simulations from the Cosmology and Astrophysics with MachinE Learning Simulations (CAMELS) project. We show that dark matter particles spread relative to their initial neighbouring distribution owing to chaotic gravitational dynamics on spatial scales comparable to their host dark matter halo. In contrast, gas in hydrodynamic simulations spreads much further from the initial neighbouring dark matter owing to feedback from supernovae (SNe) and active galactic nuclei (AGN). We show that large-scale baryon spread is very sensitive to model implementation details, with the fiducial simba model spreading ∼40 per cent of baryons >1 Mpc away compared to ∼10 per cent for the IllustrisTNG and astrid models. Increasing the efficiency of AGN-driven outflows greatly increases baryon spread while increasing the strength of SNe-driven winds can decrease spreading due to non-linear coupling of stellar and AGN feedback. We compare total matter power spectra between hydrodynamic and paired N-body simulations and demonstrate that the baryonic spread metric broadly captures the global impact of feedback on matter clustering over variations of cosmological and astrophysical parameters, initial conditions, and (to a lesser extent) galaxy formation models. Using symbolic regression, we find a function that reproduces the suppression of power by feedback as a function of wave number (k) and baryonic spread up to $k \sim 10\, h$ Mpc−1 in SIMBA while highlighting the challenge of developing models robust to variations in galaxy formation physics implementation.
We present CAMELS-ASTRID, the third suite of hydrodynamical simulations in the Cosmology and Astrophysics with MachinE Learning (CAMELS) project, along with new simulation sets that extend the model parameter space based on the previous frameworks of CAMELS-TNG and CAMELS-SIMBA, to provide broader training sets and testing grounds for machine-learning algorithms designed for cosmological studies. CAMELS-ASTRID employs the galaxy formation model following the ASTRID simulation and contains 2124 hydrodynamic simulation runs that vary three cosmological parameters (Ω
- PAR ID:
- 10479791
- Publisher / Repository:
- DOI PREFIX: 10.3847
- Date Published:
- Journal Name:
- The Astrophysical Journal
- Volume:
- 959
- Issue:
- 2
- ISSN:
- 0004-637X
- Format(s):
- Medium: X Size: Article No. 136
- Size(s):
- Article No. 136
- Sponsoring Org:
- National Science Foundation
More Like this
-
ABSTRACT -
Abstract Most diffuse baryons, including the circumgalactic medium (CGM) surrounding galaxies and the intergalactic medium (IGM) in the cosmic web, remain unmeasured and unconstrained. Fast radio bursts (FRBs) offer an unparalleled method to measure the electron dispersion measures (DMs) of ionized baryons. Their distribution can resolve the missing baryon problem and constrain the history of feedback theorized to impart significant energy to the CGM and IGM. We analyze the Cosmology and Astrophysics with Machine Learning Simulations using three suites, IllustrisTNG, SIMBA, and Astrid, each varying six parameters (two cosmological and four astrophysical feedback), for a total of 183 distinct simulation models. We find significantly different predictions between the fiducial models of the suites owing to their different implementations of feedback. SIMBA exhibits the strongest feedback, leading to the smoothest distribution of baryons and reducing the sight-line-to-sight-line variance in DMs between
z = 0 and 1. Astrid has the weakest feedback and the largest variance. We calculate FRB CGM measurements as a function of galaxy impact parameter, with SIMBA showing the weakest DMs due to aggressive active galactic nucleus (AGN) feedback and Astrid the strongest. Within each suite, the largest differences are due to varying AGN feedback. IllustrisTNG shows the most sensitivity to supernova feedback, but this is due to the change in the AGN feedback strengths, demonstrating that black holes, not stars, are most capable of redistributing baryons in the IGM and CGM. We compare our statistics directly to recent observations, paving the way for the use of FRBs to constrain the physics of galaxy formation and evolution. -
Abstract As the next generation of large galaxy surveys come online, it is becoming increasingly important to develop and understand the machine-learning tools that analyze big astronomical data. Neural networks are powerful and capable of probing deep patterns in data, but they must be trained carefully on large and representative data sets. We present a new “hump” of the Cosmology and Astrophysics with MachinE Learning Simulations (CAMELS) project: CAMELS-SAM, encompassing one thousand dark-matter-only simulations of (100
h −1cMpc)3with different cosmological parameters (Ωm andσ 8) and run through the Santa Cruz semi-analytic model for galaxy formation over a broad range of astrophysical parameters. As a proof of concept for the power of this vast suite of simulated galaxies in a large volume and broad parameter space, we probe the power of simple clustering summary statistics to marginalize over astrophysics and constrain cosmology using neural networks. We use the two-point correlation, count-in-cells, and void probability functions, and we probe nonlinear and linear scales across 0.68 <R <27h −1cMpc. We find our neural networks can both marginalize over the uncertainties in astrophysics to constrain cosmology to 3%–8% error across various types of galaxy selections, while simultaneously learning about the SC-SAM astrophysical parameters. This work encompasses vital first steps toward creating algorithms able to marginalize over the uncertainties in our galaxy formation models and measure the underlying cosmology of our Universe. CAMELS-SAM has been publicly released alongside the rest of CAMELS, and it offers great potential to many applications of machine learning in astrophysics:https://camels-sam.readthedocs.io . -
Abstract Recent work has pointed out the potential existence of a tight relation between the cosmological parameter Ω m , at fixed Ω b , and the properties of individual galaxies in state-of-the-art cosmological hydrodynamic simulations. In this paper, we investigate whether such a relation also holds for galaxies from simulations run with a different code that makes use of a distinct subgrid physics: Astrid. We also find that in this case, neural networks are able to infer the value of Ω m with a ∼10% precision from the properties of individual galaxies, while accounting for astrophysics uncertainties, as modeled in Cosmology and Astrophysics with MachinE Learning (CAMELS). This tight relationship is present at all considered redshifts, z ≤ 3, and the stellar mass, the stellar metallicity, and the maximum circular velocity are among the most important galaxy properties behind the relation. In order to use this method with real galaxies, one needs to quantify its robustness: the accuracy of the model when tested on galaxies generated by codes different from the one used for training. We quantify the robustness of the models by testing them on galaxies from four different codes: IllustrisTNG, SIMBA, Astrid, and Magneticum. We show that the models perform well on a large fraction of the galaxies, but fail dramatically on a small fraction of them. Removing these outliers significantly improves the accuracy of the models across simulation codes.more » « less
-
ABSTRACT The physics of baryons in haloes, and their subsequent influence on the total matter phase space, has a rich phenomenology and must be well understood in order to pursue a vast set of questions in both cosmology and astrophysics. We use the Cosmology and Astrophysics with MachinE Learning Simulation (Camels) suite to quantify the impact of four different galaxy formation parameters/processes (as well as two cosmological parameters) on the concentration–mass relation, cvir−Mvir. We construct a simulation-informed non-linear model for concentration as a function of halo mass, redshift, and six cosmological/astrophysical parameters. This is done for two galaxy formation models, IllustrisTNG and Simba, using 1000 simulations of each. We extract the imprints of galaxy formation across a wide range in mass $M_{\rm vir}\in [10^{11}, 10^{14.5}] \, {\rm M}_\odot \, h^{-1}$ and in redshift z ∈ [0, 6] finding many strong mass- and redshift-dependent features. Comparisons between the IllustrisTNG and Simba results show the astrophysical model choices cause significant differences in the mass and redshift dependence of these baryon imprints. Finally, we use existing observational measurements of cvir−Mvir to provide rough limits on the four astrophysical parameters. Our non-linear model is made publicly available and can be used to include Camels-based baryon imprints in any halo model-based analysis.