Predict halo masses from simulations via graph neural networks

Last update: Nov 15, 2022

Overview

HaloGraphNet

Predict halo masses from simulations via Graph Neural Networks.

Given a dark matter halo and its galaxies, creates a graph with information about the 3D position, stellar mass and other properties. Then, it trains a Graph Neural Network to predict the mass of the host halo. Data are taken from the CAMELS hydrodynamic simulations, specially suited for Machine Learning purposes. Neural nets architectures are defined making use of the package PyTorch-geometric.

See the papers arXiv:2111.08683 for more details.

Scripts

Here is a brief description of the codes included:

main.py: main driver to train and test the network.
onlytest.py: tests a pre-trained model.
hyperparams_optimization.py: optimize the hyperparameters using optuna.
camelsplots.py: plot several features of the CAMELS data.
captumtest.py: studies interpretability of the model.
halomass.py: using models trained in CAMELS, predicts the mass of real halos, such as the Milky Way and Andromeda.
visualize_graphs.py: display several halos as graphs in 2D or 3D.

The folder Hyperparameters includes files with lists of default hyperparameters, to be modified by the user. The current files contain the best values for each CAMELS simulation suite and set separately, obtained from hyperparameter optimization.

The folder Models includes some pre-trained models for the hyperparameters defined in Hyperparameters.

In the folder Source, several auxiliary routines are defined:

constants.py: basic constants and initialization.
load_data.py: contains routines to load data from simulation files.
plotting.py: includes functions for displaying the loss evolution and the results from the neural nets.
networks.py: includes the definition of the Graph Neural Networks architectures.
training.py: includes routines for training and testing the net.
galaxies.py: contains data for galaxies from the Milky Way and Andromeda halos.

Requisites

The libraries required for training the models and compute some statistics are:

numpy
pytorch-geometric
matplotlib
scipy
sklearn
optuna (only for optimization in hyperparams_optimization.py)
astropy (only for MW and M31 data in Source/galaxies.py)
captum (only for interpretability in captumtest.py)

Usage

These are some advices to employ the scripts described above:

To perform a search of the optimal hyperparameters, run hyperparams_optimization.py.
To train a model with a given set of parameters defined in params.py, run main.py.
Once a model is trained, run onlytest.py to test in the training simulation suite and cross test it in the other one included in CAMELS (IllustrisTNG and SIMBA).
Run captumtest.py to study the interpretability of the models, feature importance and saliency graphs.
Run halomass.py to infer the mass of the Milky Way and Andromeda, whose data are defined in Source/galaxies.py. For this, note that only models without the stellar mass radius as feature are considered.

Citation

If you use the code, please link this repository, and cite arXiv:2111.08683 and the DOI 10.5281/zenodo.5676528.

Contact

For comments, questions etc. you can contact me at [email protected].

Releases(v1.0)

v1.0(Apr 26, 2022)

Release version of the code.
Source code(tar.gz)
Source code(zip)

Predict halo masses from simulations via graph neural networks

Related tags

Overview

HaloGraphNet

Scripts

Requisites

Usage

Citation

Contact

You might also like...

[CIKM 2019] Code and dataset for "Fi-GNN: Modeling Feature Interactions via Graph Neural Networks for CTR Prediction"

Implementation of "GNNAutoScale: Scalable and Expressive Graph Neural Networks via Historical Embeddings" in PyTorch

Source code of NeurIPS 2021 Paper ''Be Confident! Towards Trustworthy Graph Neural Networks via Confidence Calibration''

Official Implementation of "LUNAR: Unifying Local Outlier Detection Methods via Graph Neural Networks"

My published benchmark for a Kaggle Simulations Competition

Urban mobility simulations with Python3, RLlib (Deep Reinforcement Learning) and Mesa (Agent-based modeling)

This project aims to be a handler for input creation and running of multiple RICEWQ simulations.

TUPÃ was developed to analyze electric field properties in molecular simulations

Complex-Valued Neural Networks (CVNN)Complex-Valued Neural Networks (CVNN)

Releases(v1.0)

v1.0(Apr 26, 2022)

Owner

Pablo Villanueva Domingo

Official implementation of the paper Momentum Capsule Networks (MoCapsNet)

A simple image/video to Desmos graph converter run locally

EssentialMC2 Video Understanding

Emotional conditioned music generation using transformer-based model.

Self-attentive task GAN for space domain awareness data augmentation.

Official pytorch code for SSAT: A Symmetric Semantic-Aware Transformer Network for Makeup Transfer and Removal

This is an open source library implementing hyperbox-based machine learning algorithms

Get started with Machine Learning with Python - An introduction with Python programming examples

Python script for performing depth completion from sparse depth and rgb images using the msg_chn_wacv20. model in ONNX

Transformer model implemented with Pytorch

“Robust Lightweight Facial Expression Recognition Network with Label Distribution Training”, AAAI 2021.

Deep Image Search is an AI-based image search engine that includes deep transfor learning features Extraction and tree-based vectorized search.

The Official Implementation of Neural View Synthesis and Matching for Semi-Supervised Few-Shot Learning of 3D Pose [NIPS 2021].

Library for implementing reservoir computing models (echo state networks) for multivariate time series classification and clustering.

Official implementation for (Refine Myself by Teaching Myself : Feature Refinement via Self-Knowledge Distillation, CVPR-2021)

MVS2D: Efficient Multi-view Stereo via Attention-Driven 2D Convolutions

disentanglement_lib is an open-source library for research on learning disentangled representations.

MSG-Transformer: Exchanging Local Spatial Information by Manipulating Messenger Tokens

code release for USENIX'22 paper `On the Security Risks of AutoML`

Convnet transfer - Code for paper How transferable are features in deep neural networks?