Predict halo masses from simulations via graph neural networks

Last update: Nov 15, 2022

Overview

HaloGraphNet

Predict halo masses from simulations via Graph Neural Networks.

Given a dark matter halo and its galaxies, creates a graph with information about the 3D position, stellar mass and other properties. Then, it trains a Graph Neural Network to predict the mass of the host halo. Data are taken from the CAMELS hydrodynamic simulations, specially suited for Machine Learning purposes. Neural nets architectures are defined making use of the package PyTorch-geometric.

See the papers arXiv:2111.08683 for more details.

Scripts

Here is a brief description of the codes included:

main.py: main driver to train and test the network.
onlytest.py: tests a pre-trained model.
hyperparams_optimization.py: optimize the hyperparameters using optuna.
camelsplots.py: plot several features of the CAMELS data.
captumtest.py: studies interpretability of the model.
halomass.py: using models trained in CAMELS, predicts the mass of real halos, such as the Milky Way and Andromeda.
visualize_graphs.py: display several halos as graphs in 2D or 3D.

The folder Hyperparameters includes files with lists of default hyperparameters, to be modified by the user. The current files contain the best values for each CAMELS simulation suite and set separately, obtained from hyperparameter optimization.

The folder Models includes some pre-trained models for the hyperparameters defined in Hyperparameters.

In the folder Source, several auxiliary routines are defined:

constants.py: basic constants and initialization.
load_data.py: contains routines to load data from simulation files.
plotting.py: includes functions for displaying the loss evolution and the results from the neural nets.
networks.py: includes the definition of the Graph Neural Networks architectures.
training.py: includes routines for training and testing the net.
galaxies.py: contains data for galaxies from the Milky Way and Andromeda halos.

Requisites

The libraries required for training the models and compute some statistics are:

numpy
pytorch-geometric
matplotlib
scipy
sklearn
optuna (only for optimization in hyperparams_optimization.py)
astropy (only for MW and M31 data in Source/galaxies.py)
captum (only for interpretability in captumtest.py)

Usage

These are some advices to employ the scripts described above:

To perform a search of the optimal hyperparameters, run hyperparams_optimization.py.
To train a model with a given set of parameters defined in params.py, run main.py.
Once a model is trained, run onlytest.py to test in the training simulation suite and cross test it in the other one included in CAMELS (IllustrisTNG and SIMBA).
Run captumtest.py to study the interpretability of the models, feature importance and saliency graphs.
Run halomass.py to infer the mass of the Milky Way and Andromeda, whose data are defined in Source/galaxies.py. For this, note that only models without the stellar mass radius as feature are considered.

Citation

If you use the code, please link this repository, and cite arXiv:2111.08683 and the DOI 10.5281/zenodo.5676528.

Contact

For comments, questions etc. you can contact me at [email protected].

Releases(v1.0)

v1.0(Apr 26, 2022)

Release version of the code.
Source code(tar.gz)
Source code(zip)

Predict halo masses from simulations via graph neural networks

Related tags

Overview

HaloGraphNet

Scripts

Requisites

Usage

Citation

Contact

You might also like...

[CIKM 2019] Code and dataset for "Fi-GNN: Modeling Feature Interactions via Graph Neural Networks for CTR Prediction"

Implementation of "GNNAutoScale: Scalable and Expressive Graph Neural Networks via Historical Embeddings" in PyTorch

Source code of NeurIPS 2021 Paper ''Be Confident! Towards Trustworthy Graph Neural Networks via Confidence Calibration''

Official Implementation of "LUNAR: Unifying Local Outlier Detection Methods via Graph Neural Networks"

My published benchmark for a Kaggle Simulations Competition

Urban mobility simulations with Python3, RLlib (Deep Reinforcement Learning) and Mesa (Agent-based modeling)

This project aims to be a handler for input creation and running of multiple RICEWQ simulations.

TUPÃ was developed to analyze electric field properties in molecular simulations

Complex-Valued Neural Networks (CVNN)Complex-Valued Neural Networks (CVNN)

Releases(v1.0)

v1.0(Apr 26, 2022)

Owner

Pablo Villanueva Domingo

This is implementation of AlexNet(2012) with 3D Convolution on TensorFlow (AlexNet 3D).

[ICLR'21] Counterfactual Generative Networks

DAN: Unfolding the Alternating Optimization for Blind Super Resolution

TransFGU: A Top-down Approach to Fine-Grained Unsupervised Semantic Segmentation

OMNIVORE is a single vision model for many different visual modalities

Public Models considered for emotion estimation from EEG

Visualizer using audio and semantic analysis to explore BigGAN (Brock et al., 2018) latent space.

Image-to-image regression with uncertainty quantification in PyTorch

[ICCV' 21] "Unsupervised Point Cloud Pre-training via Occlusion Completion"

Implementation of Research Paper "Learning to Enhance Low-Light Image via Zero-Reference Deep Curve Estimation"

Voice of Pajlada with model and weights.

PyTorch Implementation for Deep Metric Learning Pipelines

Colossal-AI: A Unified Deep Learning System for Large-Scale Parallel Training

Target Propagation via Regularized Inversion

Multi-query Video Retreival

Tooling for GANs in TensorFlow

The Easy-to-use Dialogue Response Selection Toolkit for Researchers

SpanNER: Named EntityRe-/Recognition as Span Prediction

Referring Video Object Segmentation

Node for thenewboston digital currency network.