Optimizers-visualized - Visualization of different optimizers on local minimas and saddle points.

Last update: Jan 01, 2022

Overview

Optimizers Visualized

Visualization of how different optimizers handle mathematical functions for optimization.

Installation
Usage
Functions for optimization
Visualization of optimizers
Links
TODO

Installation of libraries

pip install -r requirements.txt

NOTE: The optimizers used in this project are the pre-written ones in the pytorch module.

Usage

python main.py

The project is designed to be interactive, making it easy for the user to change any default values simply using stdin.

Functions for optimization

Matyas' Function

This is a relatively simple function for optimization.

Source: https://en.wikipedia.org/wiki/File:Matyas_function.pdf

Himmelblau's Function

A complex function, with multiple global minimas.

Source: https://en.wikipedia.org/wiki/File:Himmelblau_function.svg

Visualization of optimizers

All optimizers were given 100 iterations to find the global minima, from a same starting point. Learning rate was set to 0.1 for all instances, except when using SGD for minimizing Himmelblau's function.

Stochastic Gradient Descent

The vanilla stochastic gradient descent optimizer, with no additional functionalities:

theta_t = theta_t - lr * gradient

SGD on Matyas' function

We can see that SGD takes an almost direct path downwards, and then heads towards the global minima.

SGD on Himmelblau's function

SGD on Himmelblau's function fails to converge even when the learning rate is reduced from 0.1 to 0.03.

It only converges when the learning rate is further lowered to 0.01, still overshooting during the early iterations.

Root Mean Square Propagation

RMSProp with the default hyperparameters, except the learning rate.

RMSProp on Matyas' function

RMSProp first reaches a global minima in one dimension, and then switches to minimizing another dimension. This can be hurtful if there are saddle points in the function which is to be minimized.

RMSProp on Himmelblau's function

By trying to minimize one dimension first, RMSProp overshoots and has to return back to the proper path. It then minimizes the next dimension.

Adaptive Moment Estimation

Adam optimizer with the default hyperparameters, except the learning rate.

Adam on Matyas' function

Due to the momentum factor and the exponentially weighted average factor, Adam shoots past the minimal point, and returns back.

Adam on Himmelblau's function

Adam slides around the curves, again mostly due to the momentum factor.

Todos

Add more optimizers
Add more complex functions
Test out optimizers in saddle points

Optimizers-visualized - Visualization of different optimizers on local minimas and saddle points.

Related tags

Overview

Optimizers Visualized

Contents

Installation of libraries

Usage

Functions for optimization

Matyas' Function

Himmelblau's Function

Visualization of optimizers

Stochastic Gradient Descent

SGD on Matyas' function

SGD on Himmelblau's function

Root Mean Square Propagation

RMSProp on Matyas' function

RMSProp on Himmelblau's function

Adaptive Moment Estimation

Adam on Matyas' function

Adam on Himmelblau's function

Links

Todos

Owner

Gautam J

Official implementation of "Accelerating Reinforcement Learning with Learned Skill Priors", Pertsch et al., CoRL 2020

Calculates carbon footprint based on fuel mix and discharge profile at the utility selected. Can create graphs and tabular output for fuel mix based on input file of series of power drawn over a period of time.

nn_builder lets you build neural networks with less boilerplate code

DvD-TD3: Diversity via Determinants for TD3 version

Some simple programs built in Python: webcam with cv2 that detects eyes and face, with grayscale filter

The official codes of "Semi-supervised Models are Strong Unsupervised Domain Adaptation Learners".

A Low Complexity Speech Enhancement Framework for Full-Band Audio (48kHz) based on Deep Filtering.

This Repo is the official CUDA implementation of ICCV 2019 Oral paper for CARAFE: Content-Aware ReAssembly of FEatures

Ground truth data for the Optical Character Recognition of Historical Classical Commentaries.

SE3 Pose Interp - Interpolate camera pose or trajectory in SE3, pose interpolation, trajectory interpolation

Code for the paper "Controllable Video Captioning with an Exemplar Sentence"

Receptive Field Block Net for Accurate and Fast Object Detection, ECCV 2018

This project implements "virtual speed" from heart rate monito

Equivariant CNNs for the sphere and SO(3) implemented in PyTorch

Multi Camera Calibration

Fast and robust clustering of point clouds generated with a Velodyne sensor.

"SinNeRF: Training Neural Radiance Fields on Complex Scenes from a Single Image", Dejia Xu, Yifan Jiang, Peihao Wang, Zhiwen Fan, Humphrey Shi, Zhangyang Wang

PyTorch implementation of the paper Dynamic Token Normalization Improves Vision Transfromers.

This repository contains the source code for the paper First Order Motion Model for Image Animation

Light-Head R-CNN