Predict halo masses from simulations via graph neural networks

Last update: Nov 15, 2022

Overview

HaloGraphNet

Predict halo masses from simulations via Graph Neural Networks.

Given a dark matter halo and its galaxies, creates a graph with information about the 3D position, stellar mass and other properties. Then, it trains a Graph Neural Network to predict the mass of the host halo. Data are taken from the CAMELS hydrodynamic simulations, specially suited for Machine Learning purposes. Neural nets architectures are defined making use of the package PyTorch-geometric.

See the papers arXiv:2111.08683 for more details.

Scripts

Here is a brief description of the codes included:

main.py: main driver to train and test the network.
onlytest.py: tests a pre-trained model.
hyperparams_optimization.py: optimize the hyperparameters using optuna.
camelsplots.py: plot several features of the CAMELS data.
captumtest.py: studies interpretability of the model.
halomass.py: using models trained in CAMELS, predicts the mass of real halos, such as the Milky Way and Andromeda.
visualize_graphs.py: display several halos as graphs in 2D or 3D.

The folder Hyperparameters includes files with lists of default hyperparameters, to be modified by the user. The current files contain the best values for each CAMELS simulation suite and set separately, obtained from hyperparameter optimization.

The folder Models includes some pre-trained models for the hyperparameters defined in Hyperparameters.

In the folder Source, several auxiliary routines are defined:

constants.py: basic constants and initialization.
load_data.py: contains routines to load data from simulation files.
plotting.py: includes functions for displaying the loss evolution and the results from the neural nets.
networks.py: includes the definition of the Graph Neural Networks architectures.
training.py: includes routines for training and testing the net.
galaxies.py: contains data for galaxies from the Milky Way and Andromeda halos.

Requisites

The libraries required for training the models and compute some statistics are:

numpy
pytorch-geometric
matplotlib
scipy
sklearn
optuna (only for optimization in hyperparams_optimization.py)
astropy (only for MW and M31 data in Source/galaxies.py)
captum (only for interpretability in captumtest.py)

Usage

These are some advices to employ the scripts described above:

To perform a search of the optimal hyperparameters, run hyperparams_optimization.py.
To train a model with a given set of parameters defined in params.py, run main.py.
Once a model is trained, run onlytest.py to test in the training simulation suite and cross test it in the other one included in CAMELS (IllustrisTNG and SIMBA).
Run captumtest.py to study the interpretability of the models, feature importance and saliency graphs.
Run halomass.py to infer the mass of the Milky Way and Andromeda, whose data are defined in Source/galaxies.py. For this, note that only models without the stellar mass radius as feature are considered.

Citation

If you use the code, please link this repository, and cite arXiv:2111.08683 and the DOI 10.5281/zenodo.5676528.

Contact

For comments, questions etc. you can contact me at [email protected].

Releases(v1.0)

v1.0(Apr 26, 2022)

Release version of the code.
Source code(tar.gz)
Source code(zip)

Predict halo masses from simulations via graph neural networks

Related tags

Overview

HaloGraphNet

Scripts

Requisites

Usage

Citation

Contact

You might also like...

[CIKM 2019] Code and dataset for "Fi-GNN: Modeling Feature Interactions via Graph Neural Networks for CTR Prediction"

Implementation of "GNNAutoScale: Scalable and Expressive Graph Neural Networks via Historical Embeddings" in PyTorch

Source code of NeurIPS 2021 Paper ''Be Confident! Towards Trustworthy Graph Neural Networks via Confidence Calibration''

Official Implementation of "LUNAR: Unifying Local Outlier Detection Methods via Graph Neural Networks"

My published benchmark for a Kaggle Simulations Competition

Urban mobility simulations with Python3, RLlib (Deep Reinforcement Learning) and Mesa (Agent-based modeling)

This project aims to be a handler for input creation and running of multiple RICEWQ simulations.

TUPÃ was developed to analyze electric field properties in molecular simulations

Complex-Valued Neural Networks (CVNN)Complex-Valued Neural Networks (CVNN)

Releases(v1.0)

v1.0(Apr 26, 2022)

Owner

Pablo Villanueva Domingo

Real-CUGAN - Real Cascade U-Nets for Anime Image Super Resolution

Official repo for the work titled "SharinGAN: Combining Synthetic and Real Data for Unsupervised GeometryEstimation"

Repository of our paper 'Refer-it-in-RGBD' in CVPR 2021

Experiments with Fourier layers on simulation data.

No Code AI/ML platform

BOVText: A Large-Scale, Multidimensional Multilingual Dataset for Video Text Spotting

Uses Open AI Gym environment to create autonomous cryptocurrency bot to trade cryptocurrencies.

OCRA (Object-Centric Recurrent Attention) source code

Finding all things on-prem Microsoft for password spraying and enumeration.

Automatically erase objects in the video, such as logo, text, etc.

Array Camera Ptychography

Official code for UnICORNN (ICML 2021)

Train an RL agent to execute natural language instructions in a 3D Environment (PyTorch)

ByteTrack(Multi-Object Tracking by Associating Every Detection Box)のPythonでのONNX推論サンプル

Open-source python package for the extraction of Radiomics features from 2D and 3D images and binary masks.

GLANet - The code for Global and Local Alignment Networks for Unpaired Image-to-Image Translation arxiv

NAS-Bench-x11 and the Power of Learning Curves

Why Are You Weird? Infusing Interpretability in Isolation Forest for Anomaly Detection

Knowledge Management for Humans using Machine Learning & Tags

Prevent `CUDA error: out of memory` in just 1 line of code.