Skip to content

Inductive Biases in Field-Level Cosmological Inference from Galaxy Catalogs

Sep 2026 · 0 citations · 7 references
Physics Computer Science

TL;DR

Results indicate that peculiar velocities provide the dominant source of $\Omega_m$ information for set-based models in this setting, while spatial information is most effectively used by architectures that explicitly encode galaxy-galaxy relations.

Abstract

We perform field-level likelihood-free inference of the matter density parameter $\Omega_m$ from simulated galaxy catalogs using machine learning models with differing inductive biases. Using hydrodynamic simulations from CAMELS, we examine how observable choice and architecture govern cosmological information extraction. We consider galaxy positions and line-of-sight peculiar velocities, separately and jointly, and compare permutation-invariant Deep Sets, implemented with either multilayer perceptrons (MLPs) or Kolmogorov-Arnold Networks (KANs), to graph neural networks (GNNs), which explicitly encode spatial relations. We test in-distribution and out-of-distribution (OOD) performance across simulations with different subgrid galaxy-formation prescriptions. Deep Sets infer $\Omega_m$ from velocities alone with mean relative errors of approximately $18\%$ in-distribution and $\sim25\%$ OOD, with KANs and MLPs achieving comparable performance. In contrast, the same set-based approach does not yield useful $\sigma_8$ predictions in either in-distribution or cross-suite tests. Adding positions does not improve Deep Sets, while GNNs infer $\Omega_m$ with mean relative errors of about $10\%$ in-distribution and $10$--$17\%$ OOD. These results indicate that peculiar velocities provide the dominant source of $\Omega_m$ information for set-based models in this setting, while spatial information is most effectively used by architectures that explicitly encode galaxy-galaxy relations. Because the velocity inputs are exact simulated peculiar velocities, applications to survey data will require validation under realistic velocity-measurement noise, selection effects, and survey geometry.

View source

Similar papers

Review Aug 2026

FLAGS II: Constraining Galaxy Formation Models with Dimensionality Reduction of Direct Observables

Comparisons between observations of galaxies and theoretical predictions are regularly performed using physical properties, which are inferred by the often slow and biased process of SED fitting. Forward modelling facilitates a reliable alternative, whereby models are evaluated using direct observables alone. However,...

Jack C. Turner, S. Wilkins, W. Roper et al. · 0 citations
Review Aug 2026

Strong Lensing Cosmology with Population-level Calibrated Neural Ratio Estimation

This proof of concept demonstrates a potentially scalable approach for efficient cosmological parameter inference with large populations of galaxy-scale lenses observed in future surveys.

S. Jarugula, B. Nord, Aleksandra Ćiprijanović et al. · 0 citations
Review Sep 2026

Tracing the Cosmic Origins: Machine Learning Reconstruction of the Primordial Density Field from EoR Observations

Reconstructing the initial conditions of the Universe from late-time tracers would unlock cosmological information buried by non-linear structure formation and astrophysics. We reconstruct the initial density field at $z\sim300$ from simulated 21-cm and CO(1-0) line-intensity maps at $z\sim8$ generated with LIMFAST. Us...

Anchal Saxena, P. Meerburg, Guo-Chao Sun et al. · 0 citations
Review Sep 2026

A halo-based intrinsic-alignment model for simulation-based inference

Simulation-based inference and other field-level weak-lensing analyses require intrinsic-alignment (IA) models generating realistic intrinsic-ellipticity fields across two-point and non-Gaussian observables. We develop a halo-based IA prescription with separate central and satellite components and test it against intri...

M. Gatti · 0 citations
Preprint Sep 2026

Anisotropic redshift distributions in photometric galaxy clustering and their cosmological impact

Photometric galaxy clustering is a major probe of large-scale structure whose interpretation relies on accurate characterization of the galaxy redshift distribution, $n(z)$. Standard analyses assume the $n(z)$ of a selected lens sample to be isotropic, whereas observational systematics can imprint spatial variations. W...

Ze-Kang Zhang, Yun-He Wang, D. Gruen et al. · 0 citations
Review Sep 2026

Lens Modeling and Cosmological Inference from an Impure Sample of Galaxy-Galaxy Strong Lenses

The start of the Legacy Survey of Space and Time marks a new era for strong lensing science, where the number of strong lenses identified is expected to increase to $\mathcal{O}(10^5)$. In this paper we use a neural network to determine the precision with which lens parameters can be determined, using realistic simulat...

Philip Holloway, Aprajita Verma, Philip J. Marshall et al. · 0 citations

Related blog posts

MIT News · Artificial Intelligence Oct 7, 2026

Discovering the value of humanistic inquiry

Students in MIT’s Concourse program delve deeply into the human condition, debate challenging questions, and learn to develop judgment about issues that can’t be quantified.

Microsoft Research Blog Oct 7, 2026

Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses

Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.