Here I will explain the flow to deploy your custom deep learning models on Ultra96V2.

Last update: Feb 08, 2022

Overview

Xilinx_Vitis_AI

This repo will help you to Deploy your Deep Learning Model on Ultra96v2 Board.

Prerequisites

Vitis Core Development Kit 2019.2

This could be downloaded from here: Link to the websire

Vitis-AI GitHub Repository v1.1

Here is the link to the repository v1.1

Vitis-Ai Docker Container

The command to pull the container: docker pull xilinx/vitis-ai:1.1.56

XRT 2019.2

GitHub Repo Link 2019.2

Avnet Vitis Platform 2019.2

Here is the link to download the zip file Avnet Website

Ubuntu OS 18.04

Once the tools have been setup, there are five (5) main steps to targeting an AI applications to Ultra96V2 Platform:

Build the Hardware Design
Compile Your Custom Model
Build the AI Applications
Create the SD Card Content
Execute the AI Applications on hardware

Supposed that you have trained your model previously in one of the Tensorflow (.Pb), Caffe(.Caffemodel and .Prototxt) and Darknet(.Weights and .Cfg) Frameworks.

Build the Hardware Design

Clone Xilinx’s Vitis-AI github repository:

$ git clone --branch v1.1 https://github.com/Xilinx/Vitis-AI
$ cd Vitis-AI
$ export VITIS_AI_HOME = "$PWD"

Install the Avnet Vitis platform:>

Download this and extract to the hard drive of your linux machine. Then, specify the location of the Vitis platform, by creating the SDX_PLATFORM environment variable that specified to the location of the.xpfm file.

$ export SDX_PLATFORM=/home/Avnet/vitis/platform_repo/ULTRA96V2/ULTRA96V2.xpfm

Build the Hardware Project (SD Card Image)

I suggest you to download the Pre-Built from here

Compile the Trained Models

Remember that you should have pulled the docker container first.

Caffe Models:

$ cd $VITIS_AI_HOME
$ mkdir project
$ cp PATH/TO/TRAINED/MODELS  $VITIS_AI_HOME/project
$ ./docker_run.sh xilinx/vitis-ai:1.1.56
$ cd project
$ conda activate vitis-ai-caffe
$ vai_q_caffe quantize -model float.prototxt -weights float.caffemodel -calib_iter 5
$ vai_c_caffe -p .PROTOTXT -c .CAFFEMODEL -a ARCH.JSON -o OUTPUT_DIR -n NET_NAME

Tensorflow Models:

$ cd $VITIS_AI_HOME
$ mkdir project
$ cp PATH/TO/TRAINED/MODELS  $VITIS_AI_HOME/project
$ ./docker_run.sh xilinx/vitis-ai:1.1.56
$ cd project
$ conda activate vitis-ai-tensorflow
$ vai_q_tensorflow quantize --input_frozen_graph FROZEN_PB --input_nodes xxx --output_nodes yyy --input_shapes zzz --input_fn module.calib_input --calib_iter 5
$ vai_c_tensorflow -f FROZEN_PB -a ARCH.JSON -o OUTPUT_DIR -n NET_NAME

Compile the AI Application Using DNNDK APIs

The DNNDK API is the low-level API used to communicate with the AI engine (DPU). This API is the recommended API for users that will be creating their own custom neural networks.

Download and install the SDK for cross-compilation, specifying a unique and meaningful installation destination (knowing that this SDK will be specific to the Vitis-AI 1.1 DNNDK samples):

$ wget -O sdk.sh https://www.xilinx.com/bin/public/openDownload?filename=sdk.sh
$ chmod +x sdk.sh
$ ./sdk.sh -d ~/petalinux_sdk_vai_1_1_dnndk

Setup the environment for cross-compilation:

$ unset LD_LIBRARY_PATH
$ source ~/petalinux_sdk_vai_1_1_dnndk/environment-setup-aarch64-xilinx-linux

Download and extract the DNNDK runtime examples and Install the additional DNNDK runtime content:

$ wget -O vitis-ai_v1.1_dnndk.tar.gz  https://www.xilinx.com/bin/public/openDownload?filename=vitis-ai_v1.1_dnndk.tar.gz
$ tar -xvzf vitis-ai-v1.1_dnndk.tar.gz
$ cd vitis-ai-v1.1_dnndk
$ ./install.sh $SDKTARGETSYSROOT

Copy the Compiled project:

$ cp -r ../project/ .

Download and extract the additional content (images and video files) for the DNNDK examples:

$ wget -O vitis-ai_v1.1_dnndk_sample_img.tar.gz https://www.xilinx.com/bin/public/openDownload?filename=vitis-ai_v1.1_dnndk_sample_img.tar.gz
$ tar -xvzf vitis-ai_v1.1_dnndk_sample_img.tar.gz

For the custom application (project folder), create a model directory and copy the dpu_*.elf model files you previously built:

$ cd $VITIS_AI_HOME/project
$ mkdir model_for_ultra96v2
$ cp -r model_for_ultra96v2 model
$ make

NOTE: You could also edit the build.sh script to add support for the new Platforms like Ultra96V2.

Execute the AI Application on ULTRA96V2

Boot the Ultra96V2 with the pre-build sd-card image you dowloaded. For Learning How to Do This, Click HERE!
$ cd /run/media/mmcblk0p1
$ cp dpu.xclbin /usr/lib/.
Install the Vitis-AI embedded package:

$ cd runtime/vitis-ai_v1.1_dnndk 
$ source ./install.sh

Define the DISPLAY environment variable:

$ export DISPLAY=:0.0
$ xrandr --output DP-1 --mode 640x480

Run the Custom Application:

 $ cd vitis_ai_dnndk_samples
 $ ./App

Here I will explain the flow to deploy your custom deep learning models on Ultra96V2.

Related tags

Overview

Xilinx_Vitis_AI

Owner

Amin Mamandipoor

NU-Wave: A Diffusion Probabilistic Model for Neural Audio Upsampling @ INTERSPEECH 2021 Accepted

Studying Python release adoptions by looking at PyPI downloads

A Simplied Framework of GAN Inversion

Pytorch implemenation of Stochastic Multi-Label Image-to-image Translation (SMIT)

YOLOv5 detection interface - PyQt5 implementation

FCN (Fully Convolutional Network) is deep fully convolutional neural network architecture for semantic pixel-wise segmentation

This repository contains an overview of important follow-up works based on the original Vision Transformer (ViT) by Google.

🥈78th place in Riiid Answer Correctness Prediction competition

Plugin adapted from Ultralytics to bring YOLOv5 into Napari

A fast Protein Chain / Ligand Extractor and organizer.

A Comprehensive Empirical Study of Vision-Language Pre-trained Model for Supervised Cross-Modal Retrieval

Learning Representational Invariances for Data-Efficient Action Recognition

Implementations for the ICLR-2021 paper: SEED: Self-supervised Distillation For Visual Representation.

The aim of the game, as in the original one, is to find a specific image from a group of different images of a person's face

Implementation of temporal pooling methods studied in [ICIP'20] A Comparative Evaluation Of Temporal Pooling Methods For Blind Video Quality Assessment

Paper Title: Heterogeneous Knowledge Distillation for Simultaneous Infrared-Visible Image Fusion and Super-Resolution

[CVPR 2021] Forecasting the panoptic segmentation of future video frames

Implementation of Hourglass Transformer, in Pytorch, from Google and OpenAI

NudeNet: Neural Nets for Nudity Classification, Detection and selective censoring

A modern pure-Python library for reading PDF files