PyTorch implementation for the visual prior component (i.e. perception module) of the Visually Grounded Physics Learner [Li et al., 2020].

Last update: Dec 29, 2022

Related tags

Deep Learning VGPL-Visual-Prior

Overview

VGPL-Visual-Prior

PyTorch implementation for the visual prior component (i.e. perception module) of the Visually Grounded Physics Learner (VGPL). Given visual obseravtions, the visual prior proposes their corresponding particle representations, in the form of particle positions and groupings. Please see the following paper for more details.

Visual Grounding of Learned Physical Models

Yunzhu Li, Toru Lin*, Kexin Yi*, Daniel M. Bear, Daniel L. K. Yamins, Jiajun Wu, Joshua B. Tenenbaum, and Antonio Torralba

ICML 2020 [website] [paper] [video]

Demo

Input RGB videos and predictions from our learned model

Prerequisites

Python 3
PyTorch 1.0 or higher, with NVIDIA CUDA Support
Other required packages in requirements.txt

Code overview

Helper files

config.py contains all configurations used for model training, model evaluation and output generation.

dataset.py contains helper functions for loading and standardizing data and related variables. Note that paths to data directories is specified in the _DATA_DIR variable in this file, not in config.py.

loss.py contains helper functions for calculating Chamfer loss in different settings (e.g. in a single frame, across a time sequence, etc.).

model.py implements the neural network model used for prediction.

Main files

The following files can be run directly; see "Training and evaluation" section for more details.

train.py trains a model that could convert input observations into their particle representations.

eval.py evaluates a trained model by visualizing its predictions, and/or stores the output predictions in .h5 format.

Training and evaluation

Download the training and evaluation data from the following links, and put them in data folder. Optionally, download our trained model checkpoints and put them in dump folder.

MassRope [data(4.89GB)] [model]
RigidFall [data(4.87GB)] [model]

To train a model:

python train.py --set loss_type l2 dataset RigidFall

To debug (by overfitting model on small batch of data):

python train.py --set loss_type l2 dataset RigidFall debug True

To evaluate a trained model and generate outputs using our provided checkpoints:

python eval.py --set loss_type l2 dataset RigidFall n_frames 4 n_frames_eval 30 load_path dump/rigid_fall_4frame_l2.pth
python eval.py --set loss_type l2 dataset MassRope n_frames 4 n_frames_eval 30 load_path dump/mass_rope_4frame_l2.pth

See config.py for more details on customizable configurations.

Citing VGPL

If you find this codebase useful in your research, please consider citing:

@inproceedings{li2020visual,
    Title={Visual Grounding of Learned Physical Models},
    Author={Li, Yunzhu and Lin, Toru and Yi, Kexin and Bear, Daniel and Yamins, Daniel L.K. and Wu, Jiajun and Tenenbaum, Joshua B. and Torralba, Antonio},
    Booktitle={ICML},
    Year={2020}
}

@inproceedings{li2019learning,
    Title={Learning Particle Dynamics for Manipulating Rigid Bodies, Deformable Objects, and Fluids},
    Author={Li, Yunzhu and Wu, Jiajun and Tedrake, Russ and Tenenbaum, Joshua B and Torralba, Antonio},
    Booktitle={ICLR},
    Year={2019}
}

PyTorch implementation for the visual prior component (i.e. perception module) of the Visually Grounded Physics Learner [Li et al., 2020].

Related tags

Overview

VGPL-Visual-Prior

Demo

Prerequisites

Code overview

Helper files

Main files

Training and evaluation

Citing VGPL

Owner

Toru

HODEmu, is both an executable and a python library that is based on Ragagnin 2021 in prep.

This is an easy python software which allows to sort images with faces by gender and after by age.

This project uses Template Matching technique for object detecting by detection of template image over base image.

[ICML 2021] Towards Understanding and Mitigating Social Biases in Language Models

Implementation for Learning to Track with Object Permanence

I decide to sync up this repo and self-critical.pytorch. (The old master is in old master branch for archive)

Uses OpenCV and Python Code to detect a face on the screen

existing and custom freqtrade strategies supporting the new hyperstrategy format.

MINERVA: An out-of-the-box GUI tool for offline deep reinforcement learning

Python wrapper class for OpenVINO Model Server. User can submit inference request to OVMS with just a few lines of code

An AI made using artificial intelligence (AI) and machine learning algorithms (ML) .

Tensorflow implementation of MIRNet for Low-light image enhancement

A PyTorch port of the Neural 3D Mesh Renderer

CSPML (crystal structure prediction with machine learning-based element substitution)

Official repository for "Restormer: Efficient Transformer for High-Resolution Image Restoration". SOTA results for single-image motion deblurring, image deraining, image denoising (synthetic and real data), and dual-pixel defocus deblurring.

[ICCV '21] In this repository you find the code to our paper Keypoint Communities

PyTorch implementation of Convolutional Neural Fabrics http://arxiv.org/abs/1606.02492

PyTorch Personal Trainer: My framework for deep learning experiments

This folder contains the implementation of the multi-relational attribute propagation algorithm.

Implementation of ProteinBERT in Pytorch