Augmenting astrophysical scaling relations with machine learning: application to reducing the Sunyaev-Zeldovich flux-mass scatter
Digvijay Wadekar, Leander Thiele, Francisco Villaescusa-Navarro, J. Colin Hill, Miles Cranmer, David N. Spergel, Nicholas Battaglia, Daniel Anglés-Alcázar, Lars Hernquist, Shirley Ho
Code Available — Be the first to reproduce this paper.
ReproduceCode
- github.com/jaywadekar/scalingrelations_mlOfficialIn papernone★ 2
Abstract
Complex astrophysical systems often exhibit low-scatter relations between observable properties (e.g., luminosity, velocity dispersion, oscillation period). These scaling relations illuminate the underlying physics, and can provide observational tools for estimating masses and distances. Machine learning can provide a fast and systematic way to search for new scaling relations (or for simple extensions to existing relations) in abstract high-dimensional parameter spaces. We use a machine learning tool called symbolic regression (SR), which models patterns in a dataset in the form of analytic equations. We focus on the Sunyaev-Zeldovich flux-cluster mass relation (Y_SZ-M), the scatter in which affects inference of cosmological parameters from cluster abundance data. Using SR on the data from the IllustrisTNG hydrodynamical simulation, we find a new proxy for cluster mass which combines Y_SZ and concentration of ionized gas (c_gas): M Y_conc^3/5 Y_SZ^3/5 (1-A\, c_gas). Y_conc reduces the scatter in the predicted M by 20-30\% for large clusters (M 10^14\, h^-1 \, M_), as compared to using just Y_SZ. We show that the dependence on c_gas is linked to cores of clusters exhibiting larger scatter than their outskirts. Finally, we test Y_conc on clusters from CAMELS simulations and show that Y_conc is robust against variations in cosmology, subgrid physics, and cosmic variance. Our results and methodology can be useful for accurate multiwavelength cluster mass estimation from upcoming CMB and X-ray surveys like ACT, SO, eROSITA and CMB-S4.