My implementation of transformers related papers for computer vision in pytorch

Last update: Nov 10, 2021

Overview

vision_transformers

This is my personnal repo to implement new transofrmers based and other computer vision DL models

I am currenlty working without a lot of GPU ressources therefore I mainly trained models on CIFAR 10. But my implementation are build to be fast and effective at scale.

Current paper implemented:

An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale, from Dosovitskiy et al (2020)
Patch Are All You Need ? anonymous

Baseline:

Deep Residual Learning for Image Recognition, from He et al (2015)

Models are implemented in pure pytorch and trained via pytorchlightning. Dependencies are managed by poetry. It is included an Dockerfile to create a cuda ready container with jupyter lab inside. On the development part, I use jupytext in order to avoid commit every metadata change on the notebook. Fully tested with pytest and formatted with black and isort.

If you want to create a project with similar config, just use my boilerplat.

How to use it ?

first install the dependecies:

poetry install

Then, only for development:

add the precommit hook

poetry run pre-commit install

sync the notebook (only once)

poetry shell
make notebook-sync

launch a jupyter lab session

poetry run jupyter lab

Use tensorboard

poetry shell
make tensorboard

Format the code without the precommit hook

poetry shell
make formatting

Tests:

to run the tests:

poetry shell
make tests

You might also like...

Build fully-functioning computer vision models with PyTorch

Detecto is a Python package that allows you to build fully-functioning computer vision and object detection models with just 5 lines of code. Inferenc

576 Dec 29, 2022

A PyTorch-Based Framework for Deep Learning in Computer Vision

TorchCV: A PyTorch-Based Framework for Deep Learning in Computer Vision @misc{you2019torchcv, author = {Ansheng You and Xiangtai Li and Zhen Zhu a

2.2k Jan 9, 2023

Open Source Differentiable Computer Vision Library for PyTorch

Kornia is a differentiable computer vision library for PyTorch. It consists of a set of routines and differentiable modules to solve generic computer

7.6k Jan 4, 2023

An Agnostic Computer Vision Framework - Pluggable to any Training Library: Fastai, Pytorch-Lightning with more to come

IceVision is the first agnostic computer vision framework to offer a curated collection with hundreds of high-quality pre-trained models from torchvision, MMLabs, and soon Pytorch Image Models. It orchestrates the end-to-end deep learning workflow allowing to train networks with easy-to-use robust high-performance libraries such as Pytorch-Lightning and Fastai

789 Dec 29, 2022

My implementation of transformers related papers for computer vision in pytorch

Related tags

Overview

vision_transformers

How to use it ?

launch a jupyter lab session

Use tensorboard

Format the code without the precommit hook

Tests:

You might also like...

Build fully-functioning computer vision models with PyTorch

A PyTorch-Based Framework for Deep Learning in Computer Vision

Open Source Differentiable Computer Vision Library for PyTorch

An Agnostic Computer Vision Framework - Pluggable to any Training Library: Fastai, Pytorch-Lightning with more to come

Spiking Neural Network for Computer Vision using SpikingJelly framework and Pytorch-Lightning

Implementation of self-attention mechanisms for general purpose. Focused on computer vision modules. Ongoing repository.

The Incredible PyTorch: a curated list of tutorials, papers, projects, communities and more relating to PyTorch.

Explainability for Vision Transformers (in PyTorch)

PyTorch code for Vision Transformers training with the Self-Supervised learning method DINO

Releases(0.1.0)

0.1.0(Nov 10, 2021)

Owner

samsja

A framework to train language models to learn invariant representations.

A Novel Plug-in Module for Fine-grained Visual Classification

Code for ICMI2020 and ICMI2021 papers: "Studying Person-Specific Pointing and Gaze Behavior for Multimodal Referencing of Outside Objects from a Moving Vehicle" and "ML-PersRef: A Machine Learning-based Personalized Multimodal Fusion Approach for Referencing Outside Objects From a Moving Vehicle"

Convenient tool for speeding up the intern/officer review process.

Alternatives to Deep Neural Networks for Function Approximations in Finance

Original code for "Zero-Shot Domain Adaptation with a Physics Prior"

Deep Reinforcement Learning for Multiplayer Online Battle Arena

Binary Passage Retriever (BPR) - an efficient passage retriever for open-domain question answering

Transformer in Vision

Machine Unlearning with SISA

Code for the Active Speakers in Context Paper (CVPR2020)

The Balloon Learning Environment - flying stratospheric balloons with deep reinforcement learning.

NudeNet: Neural Nets for Nudity Classification, Detection and selective censoring

A Python Automated Machine Learning tool that optimizes machine learning pipelines using genetic programming.

salabim - discrete event simulation in Python

Lane assist for ETS2, built with the ultra-fast-lane-detection model.

Direct design of biquad filter cascades with deep learning by sampling random polynomials.

Designing a Practical Degradation Model for Deep Blind Image Super-Resolution (ICCV, 2021) (PyTorch) - We released the training code!

Cortex-compatible model server for Python and TensorFlow

The repository offers the official implementation of our paper in PyTorch.