[CVPR 2016] Unsupervised Feature Learning by Image Inpainting using GANs

Last update: Dec 31, 2022

Overview

Context Encoders: Feature Learning by Inpainting

CVPR 2016

[Project Website] [Imagenet Results]

Sample results on held-out images:

This is the training code for our CVPR 2016 paper on Context Encoders for learning deep feature representation in an unsupervised manner by image inpainting. Context Encoders are trained jointly with reconstruction and adversarial loss. This repo contains quick demo, training/testing code for center region inpainting and training/testing code for arbitray random region inpainting. This code is adapted from an initial fork of Soumith's DCGAN implementation. Scroll down to try out a quick demo or train your own inpainting models!

If you find Context Encoders useful in your research, please cite:

@inproceedings{pathakCVPR16context,
    Author = {Pathak, Deepak and Kr\"ahenb\"uhl, Philipp and Donahue, Jeff and Darrell, Trevor and Efros, Alexei},
    Title = {Context Encoders: Feature Learning by Inpainting},
    Booktitle = {Computer Vision and Pattern Recognition ({CVPR})},
    Year = {2016}
}

Semantic Inpainting Demo
Train Context Encoders
Download Features Caffemodel
TensorFlow Implementation
Project Website
Download Dataset

1) Semantic Inpainting Demo

Install Torch: http://torch.ch/docs/getting-started.html#_
Clone the repository

git clone https://github.com/pathak22/context-encoder.git

Demo

cd context-encoder
bash ./models/scripts/download_inpaintCenter_models.sh
# This will populate the `./models/` folder with trained models.

net=models/inpaintCenter/paris_inpaintCenter.t7 name=paris_result imDir=images/paris overlapPred=4 manualSeed=222 batchSize=21 gpu=1 th demo.lua
net=models/inpaintCenter/imagenet_inpaintCenter.t7 name=imagenet_result imDir=images/imagenet overlapPred=4 manualSeed=222 batchSize=21 gpu=1 th demo.lua
net=models/inpaintCenter/paris_inpaintCenter.t7 name=ucberkeley_result imDir=images/ucberkeley overlapPred=4 manualSeed=222 batchSize=4 gpu=1 th demo.lua
# Note: If you are running on cpu, use gpu=0
# Note: samples given in ./images/* are held-out images

2) Train Context Encoders

If you could successfully run the above demo, run following steps to train your own context encoder model for image inpainting.

[Optional] Install Display Package as follows. If you don't want to install it, then set display=0 in train.lua.

luarocks install https://raw.githubusercontent.com/szym/display/master/display-scm-0.rockspec
cd ~
th -ldisplay.start 8000
# if working on server machine create tunnel: ssh -f -L 8000:localhost:8000 -N server_address.com
# on client side, open in browser: http://localhost:8000/

Make the dataset folder.

mkdir -p /path_to_wherever_you_want/mydataset/train/images/
# put all training images inside mydataset/train/images/
mkdir -p /path_to_wherever_you_want/mydataset/val/images/
# put all val images inside mydataset/val/images/
cd context-encoder/
ln -sf /path_to_wherever_you_want/mydataset dataset

Train the model

# For training center region inpainting model, run:
DATA_ROOT=dataset/train display_id=11 name=inpaintCenter overlapPred=4 wtl2=0.999 nBottleneck=4000 niter=500 loadSize=350 fineSize=128 gpu=1 th train.lua

# For training random region inpainting model, run:
DATA_ROOT=dataset/train display_id=11 name=inpaintRandomNoOverlap useOverlapPred=0 wtl2=0.999 nBottleneck=4000 niter=500 loadSize=350 fineSize=128 gpu=1 th train_random.lua
# or use fineSize=64 to train to generate 64x64 sized image (results are better):
DATA_ROOT=dataset/train display_id=11 name=inpaintRandomNoOverlap useOverlapPred=0 wtl2=0.999 nBottleneck=4000 niter=500 loadSize=350 fineSize=64 gpu=1 th train_random.lua

Test the model

# For training center region inpainting model, run:
DATA_ROOT=dataset/val net=checkpoints/inpaintCenter_500_net_G.t7 name=test_patch overlapPred=4 manualSeed=222 batchSize=30 loadSize=350 gpu=1 th test.lua
DATA_ROOT=dataset/val net=checkpoints/inpaintCenter_500_net_G.t7 name=test_full overlapPred=4 manualSeed=222 batchSize=30 loadSize=129 gpu=1 th test.lua

# For testing random region inpainting model, run (with fineSize=64 or 124, same as training):
DATA_ROOT=dataset/val net=checkpoints/inpaintRandomNoOverlap_500_net_G.t7 name=test_patch_random useOverlapPred=0 manualSeed=222 batchSize=30 loadSize=350 gpu=1 th test_random.lua
DATA_ROOT=dataset/val net=checkpoints/inpaintRandomNoOverlap_500_net_G.t7 name=test_full_random useOverlapPred=0 manualSeed=222 batchSize=30 loadSize=129 gpu=1 th test_random.lua

3) Download Features Caffemodel

Features for context encoder trained with reconstruction loss.

4) TensorFlow Implementation

Checkout the TensorFlow implementation of our paper by Taeksoo here. However, it does not implement full functionalities of our paper.

5) Project Website

Click here.

6) Paris Street-View Dataset

Please email me if you need the dataset and I will share a private link with you. I can't post the public link to this dataset due to the policy restrictions from Google Street View.

[CVPR 2016] Unsupervised Feature Learning by Image Inpainting using GANs

Related tags

Overview

Context Encoders: Feature Learning by Inpainting

CVPR 2016

[Project Website] [Imagenet Results]

Contents

1) Semantic Inpainting Demo

2) Train Context Encoders

3) Download Features Caffemodel

4) TensorFlow Implementation

5) Project Website

6) Paris Street-View Dataset

Owner

Deepak Pathak

External Attention Network

PyTorch implementation of our paper How robust are discriminatively trained zero-shot learning models?

A PyTorch Implementation of Gated Graph Sequence Neural Networks (GGNN)

an implementation of 3D Ken Burns Effect from a Single Image using PyTorch

KIND: an Italian Multi-Domain Dataset for Named Entity Recognition

Analysing poker data from home games with friends

Deep learning-based approach to discovering Granger causality networks in multivariate time series

SalFBNet: Learning Pseudo-Saliency Distribution via Feedback Convolutional Networks

CVPR2022 paper "Dense Learning based Semi-Supervised Object Detection"

Pytorch implementation for Semantic Segmentation/Scene Parsing on MIT ADE20K dataset

AI-generated-characters for Learning and Wellbeing

BigDetection: A Large-scale Benchmark for Improved Object Detector Pre-training

The official repo of the CVPR2021 oral paper: Representative Batch Normalization with Feature Calibration

Bootstrapped Representation Learning on Graphs

A hue shift helper for OBS

Enhancing Knowledge Tracing via Adversarial Training

This repository contains answers of the Shopify Summer 2022 Data Science Intern Challenge.

[CVPR-2021] UnrealPerson: An adaptive pipeline for costless person re-identification

Puzzle-CAM: Improved localization via matching partial and full features.

Determined: Deep Learning Training Platform