A higher performance pytorch implementation of DeepLab V3 Plus(DeepLab v3+)

Last update: Nov 22, 2022

Related tags

Overview

A Higher Performance Pytorch Implementation of DeepLab V3 Plus

Introduction

This repo is an (re-)implementation of Encoder-Decoder with Atrous Separable Convolution for Semantic Image Segmentation in PyTorch for semantic image segmentation on the PASCAL VOC dataset. And this repo has a higher mIoU of 79.19% than the result of paper which is 78.85%.

Requirements

Python(3.6) and Pytorch(0.4.1) is necessary before running the scripts. To install the required python packages(expect PyTorch), run

pip install -r requirements.txt

Datasets

To train and validate the network, this repo use the augmented PASCAL VOC 2012 dataset which contains 10582 images for training and 1449 images for validation. To use the dataset, you can download the PASCAL VOC training/validation data (2GB tar file) here and download the SegmentationClassAug from dropbox or Baidu Netdisk

Training

Before training, you should clone this repo:

git clone git@github.com:hualin95/Deeplab-v3plus.git

You can begin training by running the train.py.

#training
cd Deeplab-v3plus-master/tools/   
python train.py

You are expected to achieve PA:94.77%, MPA:88.48%, MIoU:79.19%, FWIoU:90.53% on the validation.

#Monitoring
tensorboard --logdir=runs/ --port=80

Performance

VOC2012: after 30k iterations with a batch size of 16.

Backbone	train OS	eval OS	MS	mIoU paper	mIoU repo
Resnet101	16	16	No	78.85%	79.19%

TODO

Resnet as Network Backbone
Implement depthwise separable convolutions
Multi-GPU support
Model pretrained on MS-COCO
Xception as Network Backbone

A higher performance pytorch implementation of DeepLab V3 Plus(DeepLab v3+)

Related tags

Overview

A Higher Performance Pytorch Implementation of DeepLab V3 Plus

Introduction

Requirements

Datasets

Training

Performance

TODO

Owner

linhua

Implementation of Perceiver, General Perception with Iterative Attention, in Pytorch

Pytorch implementation of "Grad-TTS: A Diffusion Probabilistic Model for Text-to-Speech"

Rank 1st in the public leaderboard of ScanRefer (2021-03-18)

A Flow-based Generative Network for Speech Synthesis

Deep Anomaly Detection with Outlier Exposure (ICLR 2019)

Event queue (Equeue) dialect is an MLIR Dialect that models concurrent devices in terms of control and structure.

PyTorch implementation of PP-LCNet

Dynamic Realtime Animation Control

SlideGraph+: Whole Slide Image Level Graphs to Predict HER2 Status in Breast Cancer

Pytorch implementation of Learning with Opponent-Learning Awareness

Iowa Project - My second project done at General Assembly, focused on feature engineering and understanding Linear Regression as a concept

Pytorch tutorials for Neural Style transfert

Music library streaming app written in Flask & VueJS

Make your AirPlay devices as TTS speakers

Code repository for "Reducing Underflow in Mixed Precision Training by Gradient Scaling" presented at IJCAI '20

Pytorch implementation for "Adversarial Robustness under Long-Tailed Distribution" (CVPR 2021 Oral)

[CVPR2022] Representation Compensation Networks for Continual Semantic Segmentation

Single Image Random Dot Stereogram for Tensorflow

Implementation of ConvMixer for "Patches Are All You Need? 🤷"

Learning View Priors for Single-view 3D Reconstruction (CVPR 2019)