-
Fast Forward Modelling of Galaxy Spatial and Statistical Distributions
Authors:
Pascale Berner,
Alexandre Refregier,
Beatrice Moser,
Luca Tortorelli,
Luis Fernando Machado Poletti Valle,
Tomasz Kacprzak
Abstract:
A forward modelling approach provides simple, fast and realistic simulations of galaxy surveys, without a complex underlying model. For this purpose, galaxy clustering needs to be simulated accurately, both for the usage of clustering as its own probe and to control systematics. We present a forward model to simulate galaxy surveys, where we extend the Ultra-Fast Image Generator to include galaxy…
▽ More
A forward modelling approach provides simple, fast and realistic simulations of galaxy surveys, without a complex underlying model. For this purpose, galaxy clustering needs to be simulated accurately, both for the usage of clustering as its own probe and to control systematics. We present a forward model to simulate galaxy surveys, where we extend the Ultra-Fast Image Generator to include galaxy clustering. We use the distribution functions of the galaxy properties, derived from a forward model adjusted to observations. This population model jointly describes the luminosity functions, sizes, ellipticities, SEDs and apparent magnitudes. To simulate the positions of galaxies, we then use a two-parameter relation between galaxies and halos with Subhalo Abundance Matching (SHAM). We simulate the halos and subhalos using the fast PINOCCHIO code, and a method to extract the surviving subhalos from the merger history. Our simulations contain a red and a blue galaxy population, for which we build a SHAM model based on star formation quenching. For central galaxies, mass quenching is controlled with the parameter M$_{\mathrm{limit}}$, with blue galaxies residing in smaller halos. For satellite galaxies, environmental quenching is implemented with the parameter t$_{\mathrm{quench}}$, where blue galaxies occupy only recently merged subhalos. We build and test our model by comparing to imaging data from the Dark Energy Survey Year 1. To ensure completeness in our simulations, we consider the brightest galaxies with $i<20$. We find statistical agreement between our simulations and the data for two-point correlation functions on medium to large scales. Our model provides constraints on the two SHAM parameters M$_{\mathrm{limit}}$ and t$_{\mathrm{quench}}$ and offers great prospects for the quick generation of galaxy mock catalogues, optimized to agree with observations.
△ Less
Submitted 24 February, 2024; v1 submitted 23 October, 2023;
originally announced October 2023.
-
$\mathbf{12\times2}$pt combined probes: pipeline, neutrino mass, and data compression
Authors:
Alexander Reeves,
Andrina Nicola,
Alexandre Refregier,
Tomasz Kacprzak,
Luis Fernando Machado Poletti Valle
Abstract:
With the rapid advance of wide-field surveys it is increasingly important to perform combined cosmological probe analyses. We present a new pipeline for simulation-based multi-probe analyses, which combines tomographic large-scale structure (LSS) probes (weak lensing and galaxy clustering) with cosmic microwave background (CMB) primary and lensing data. These are combined at the $C_\ell$-level, yi…
▽ More
With the rapid advance of wide-field surveys it is increasingly important to perform combined cosmological probe analyses. We present a new pipeline for simulation-based multi-probe analyses, which combines tomographic large-scale structure (LSS) probes (weak lensing and galaxy clustering) with cosmic microwave background (CMB) primary and lensing data. These are combined at the $C_\ell$-level, yielding 12 distinct auto- and cross-correlations. The pipeline is based on $\texttt{UFalconv2}$, a framework to generate fast, self-consistent map-level realizations of cosmological probes from input lightcones, which is applied to the $\texttt{CosmoGridV1}$ N-body simulation suite. It includes a non-Gaussian simulation-based covariance for the LSS tracers, several data compression schemes, and a neural network emulator for accelerated theoretical predictions. We validate our framework, apply it to a simulated $12\times2$pt tomographic analysis of KiDS, BOSS, and $\textit{Planck}$, and forecast constraints for a $Λ$CDM model with a variable neutrino mass. We find that, while the neutrino mass constraints are driven by the CMB data, the addition of LSS data helps to break degeneracies and improves the constraint by up to 35%. For a fiducial $M_ν=0.15\mathrm{eV}$, a full combination of the above CMB+LSS data would enable a $3σ$ constraint on the neutrino mass. We explore data compression schemes and find that MOPED outperforms PCA. We also study the impact of an internal lensing tension in the CMB data, parametrized by $A_L$, on the neutrino mass constraint, finding that the addition of LSS to CMB data including all cross-correlations is able to mitigate the impact of this systematic. $\texttt{UFalconv2}$ and a MOPED compressed $\textit{Planck}$ CMB primary + CMB lensing likelihood are made publicly available. [abridged]
△ Less
Submitted 6 December, 2023; v1 submitted 6 September, 2023;
originally announced September 2023.
-
The Circumgalactic Medium from the CAMELS Simulations: Forecasting Constraints on Feedback Processes from Future Sunyaev-Zeldovich Observations
Authors:
Emily Moser,
Nicholas Battaglia,
Daisuke Nagai,
Erwin Lau,
Luis Fernando Machado Poletti Valle,
Francisco Villaescusa-Navarro,
Stefania Amodeo,
Daniel Angles-Alcazar,
Greg L. Bryan,
Romeel Dave,
Lars Hernquist,
Mark Vogelsberger
Abstract:
The cycle of baryons through the circumgalactic medium (CGM) is important to understand in the context of galaxy formation and evolution. In this study we forecast constraints on the feedback processes heating the CGM with current and future Sunyaev-Zeldovich (SZ) observations. To constrain these processes, we use a suite of cosmological simulations, the Cosmology and Astrophysics with MachinE Lea…
▽ More
The cycle of baryons through the circumgalactic medium (CGM) is important to understand in the context of galaxy formation and evolution. In this study we forecast constraints on the feedback processes heating the CGM with current and future Sunyaev-Zeldovich (SZ) observations. To constrain these processes, we use a suite of cosmological simulations, the Cosmology and Astrophysics with MachinE Learning Simulations (CAMELS), that varies four different feedback parameters of two previously existing hydrodynamical simulations, IllustrisTNG and SIMBA. We capture the dependencies of SZ radial profiles on these feedback parameters with an emulator, calculate their derivatives, and forecast future constraints on these feedback parameters from upcoming experiments. We find that for a DESI-like (Dark Energy Spectroscopic Instrument) galaxy sample observed by the Simons Observatory all four feedback parameters are able to be constrained (some within the $10\%$ level), indicating that future observations will be able to further restrict the parameter space for these sub-grid models. Given the modeled galaxy sample and forecasted errors in this work, we find that the inner SZ profiles contribute more to the constraining power than the outer profiles. Finally, we find that, despite the wide range of AGN feedback parameter variation in the CAMELS simulation suite, we cannot reproduce the tSZ signal of galaxies selected by the Baryon Oscillation Spectroscopic Survey as measured by the Atacama Cosmology Telescope.
△ Less
Submitted 7 January, 2022;
originally announced January 2022.
-
The CAMELS project: public data release
Authors:
Francisco Villaescusa-Navarro,
Shy Genel,
Daniel Anglés-Alcázar,
Lucia A. Perez,
Pablo Villanueva-Domingo,
Digvijay Wadekar,
Helen Shao,
Faizan G. Mohammad,
Sultan Hassan,
Emily Moser,
Erwin T. Lau,
Luis Fernando Machado Poletti Valle,
Andrina Nicola,
Leander Thiele,
Yongseok Jo,
Oliver H. E. Philcox,
Benjamin D. Oppenheimer,
Megan Tillman,
ChangHoon Hahn,
Neerav Kaushal,
Alice Pisani,
Matthew Gebhardt,
Ana Maria Delgado,
Joyce Caliendo,
Christina Kreisch
, et al. (22 additional authors not shown)
Abstract:
The Cosmology and Astrophysics with MachinE Learning Simulations (CAMELS) project was developed to combine cosmology with astrophysics through thousands of cosmological hydrodynamic simulations and machine learning. CAMELS contains 4,233 cosmological simulations, 2,049 N-body and 2,184 state-of-the-art hydrodynamic simulations that sample a vast volume in parameter space. In this paper we present…
▽ More
The Cosmology and Astrophysics with MachinE Learning Simulations (CAMELS) project was developed to combine cosmology with astrophysics through thousands of cosmological hydrodynamic simulations and machine learning. CAMELS contains 4,233 cosmological simulations, 2,049 N-body and 2,184 state-of-the-art hydrodynamic simulations that sample a vast volume in parameter space. In this paper we present the CAMELS public data release, describing the characteristics of the CAMELS simulations and a variety of data products generated from them, including halo, subhalo, galaxy, and void catalogues, power spectra, bispectra, Lyman-$α$ spectra, probability distribution functions, halo radial profiles, and X-rays photon lists. We also release over one thousand catalogues that contain billions of galaxies from CAMELS-SAM: a large collection of N-body simulations that have been combined with the Santa Cruz Semi-Analytic Model. We release all the data, comprising more than 350 terabytes and containing 143,922 snapshots, millions of halos, galaxies and summary statistics. We provide further technical details on how to access, download, read, and process the data at \url{https://camels.readthedocs.io}.
△ Less
Submitted 4 January, 2022;
originally announced January 2022.
-
The CAMELS Multifield Dataset: Learning the Universe's Fundamental Parameters with Artificial Intelligence
Authors:
Francisco Villaescusa-Navarro,
Shy Genel,
Daniel Angles-Alcazar,
Leander Thiele,
Romeel Dave,
Desika Narayanan,
Andrina Nicola,
Yin Li,
Pablo Villanueva-Domingo,
Benjamin Wandelt,
David N. Spergel,
Rachel S. Somerville,
Jose Manuel Zorrilla Matilla,
Faizan G. Mohammad,
Sultan Hassan,
Helen Shao,
Digvijay Wadekar,
Michael Eickenberg,
Kaze W. K. Wong,
Gabriella Contardo,
Yongseok Jo,
Emily Moser,
Erwin T. Lau,
Luis Fernando Machado Poletti Valle,
Lucia A. Perez
, et al. (3 additional authors not shown)
Abstract:
We present the Cosmology and Astrophysics with MachinE Learning Simulations (CAMELS) Multifield Dataset, CMD, a collection of hundreds of thousands of 2D maps and 3D grids containing many different properties of cosmic gas, dark matter, and stars from 2,000 distinct simulated universes at several cosmic times. The 2D maps and 3D grids represent cosmic regions that span $\sim$100 million light year…
▽ More
We present the Cosmology and Astrophysics with MachinE Learning Simulations (CAMELS) Multifield Dataset, CMD, a collection of hundreds of thousands of 2D maps and 3D grids containing many different properties of cosmic gas, dark matter, and stars from 2,000 distinct simulated universes at several cosmic times. The 2D maps and 3D grids represent cosmic regions that span $\sim$100 million light years and have been generated from thousands of state-of-the-art hydrodynamic and gravity-only N-body simulations from the CAMELS project. Designed to train machine learning models, CMD is the largest dataset of its kind containing more than 70 Terabytes of data. In this paper we describe CMD in detail and outline a few of its applications. We focus our attention on one such task, parameter inference, formulating the problems we face as a challenge to the community. We release all data and provide further technical details at https://camels-multifield-dataset.readthedocs.io.
△ Less
Submitted 22 September, 2021;
originally announced September 2021.
-
SHA** the Gas: Understanding Gas Shapes in Dark Matter Haloes with Interpretable Machine Learning
Authors:
Luis Fernando Machado Poletti Valle,
Camille Avestruz,
David J. Barnes,
Arya Farahi,
Erwin T. Lau,
Daisuke Nagai
Abstract:
The non-spherical shapes of dark matter and gas distributions introduce systematic uncertainties that affect observable-mass relations and selection functions of galaxy groups and clusters. However, the triaxial gas distributions depend on the non-linear physical processes of halo formation histories and baryonic physics, which are challenging to model accurately. In this study we explore a machin…
▽ More
The non-spherical shapes of dark matter and gas distributions introduce systematic uncertainties that affect observable-mass relations and selection functions of galaxy groups and clusters. However, the triaxial gas distributions depend on the non-linear physical processes of halo formation histories and baryonic physics, which are challenging to model accurately. In this study we explore a machine learning approach for modelling the dependence of gas shapes on dark matter and baryonic properties. With data from the IllustrisTNG hydrodynamical cosmological simulations, we develop a machine learning pipeline that applies \pkg{XGBoost}, an implementation of gradient boosted decision trees, to predict radial profiles of gas shapes from halo properties. We show that \pkg{XGBoost} models can accurately predict gas shape profiles in dark matter haloes. We also explore model interpretability with \pkg{SHAP}, a method that identifies the most predictive properties at different halo radii. We find that baryonic properties best predict gas shapes in halo cores, whereas dark matter shapes are the main predictors in the halo outskirts. This work demonstrates the power of interpretable machine learning in modelling observable properties of dark matter haloes in the era of multi-wavelength cosmological surveys.
△ Less
Submitted 2 August, 2021; v1 submitted 25 November, 2020;
originally announced November 2020.