DeepLab2: A TensorFlow Library for Deep Labeling

Last update: Jan 04, 2023

Related tags

Overview

DeepLab2: A TensorFlow Library for Deep Labeling

DeepLab2 is a TensorFlow library for deep labeling, aiming to provide a unified and state-of-the-art TensorFlow codebase for dense pixel labeling tasks, including, but not limited to semantic segmentation, instance segmentation, panoptic segmentation, depth estimation, or even video panoptic segmentation.

Deep labeling refers to solving computer vision problems by assigning a predicted value for each pixel in an image with a deep neural network. As long as the problem of interest could be formulated in this way, DeepLab2 should serve the purpose. Additionally, this codebase includes our recent and state-of-the-art research models on deep labeling. We hope you will find it useful for your projects.

Installation

See Installation.

Dataset preparation

The dataset needs to be converted to TFRecord. We provide some examples below.

Some guidances about how to convert your own dataset.

Your Own Dataset

Projects

We list a few projects that use DeepLab2.

Colab Demo

Colab notebook for off-the-shelf inference.

Running DeepLab2

See Getting Started. In short, run the following command:

To run DeepLab2 on GPUs, the following command should be used:

python training/train.py \
    --config_file=${CONFIG_FILE} \
    --mode={train | eval | train_and_eval | continuous_eval} \
    --model_dir=${BASE_MODEL_DIRECTORY} \
    --num_gpus=${NUM_GPUS}

Change logs

See Change logs for recent updates.

Contacts (Maintainers)

Please check FAQ if you have some questions before reporting the issues.

Mark Weber, github: markweberdev
Huiyu Wang, github: csrhddlam
Siyuan Qiao, github: joe-siyuan-qiao
Jun Xie, github: clairexie
Maxwell D. Collins, github: mcollinswisc
YuKun Zhu, github: yknzhu
Liangzhe Yuan, github: yuanliangzhe
Dahun Kim, github: mcahny
Qihang Yu, github: yucornetto
Liang-Chieh Chen, github: aquariusjay

Disclaimer

Note that this library contains our re-implemented DeepLab models in TensorFlow2, and thus may have some minor differences from the published papers (e.g., learning rate).
This is not an official Google product.

Citing DeepLab2

If you find DeepLab2 useful for your project, please consider citing DeepLab2 along with the relevant DeepLab series.

DeepLab2:

@article{deeplab2_2021,
  author={Mark Weber and Huiyu Wang and Siyuan Qiao and Jun Xie and Maxwell D. Collins and Yukun Zhu and Liangzhe Yuan and Dahun Kim and Qihang Yu and Daniel Cremers and Laura Leal-Taixe and Alan L. Yuille and Florian Schroff and Hartwig Adam and Liang-Chieh Chen},
  title={{DeepLab2: A TensorFlow Library for Deep Labeling}},
  journal={arXiv: 2106.09748},
  year={2021}
}

References

Marius Cordts, Mohamed Omran, Sebastian Ramos, Timo Rehfeld, Markus Enzweiler, Rodrigo Benenson, Uwe Franke, Stefan Roth, and Bernt Schiele. "The cityscapes dataset for semantic urban scene understanding." In CVPR, 2016.
Andreas Geiger, Philip Lenz, and Raquel Urtasun. "Are we ready for autonomous driving? the kitti vision benchmark suite." In CVPR, 2012.
Jens Behley, Martin Garbade, Andres Milioto, Jan Quenzel, Sven Behnke, Cyrill Stachniss, and Jurgen Gall. "Semantickitti: A dataset for semantic scene understanding of lidar sequences." In ICCV, 2019.
Alexander Kirillov, Kaiming He, Ross Girshick, Carsten Rother, and Piotr Dollar. "Panoptic segmentation." In CVPR, 2019.
Dahun Kim, Sanghyun Woo, Joon-Young Lee, and In So Kweon. "Video panoptic segmentation." In CVPR, 2020.
Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollar, and C Lawrence Zitnick. "Microsoft COCO: Common objects in context." In ECCV, 2014.
Patrick Dendorfer, Aljosa Osep, Anton Milan, Konrad Schindler, Daniel Cremers, Ian Reid, Stefan Roth, and Laura Leal-Taixe. "MOTChallenge: A Benchmark for Single-camera Multiple Target Tracking." IJCV, 2020.

DeepLab2: A TensorFlow Library for Deep Labeling

Related tags

Overview

DeepLab2: A TensorFlow Library for Deep Labeling

Installation

Dataset preparation

Projects

Colab Demo

Running DeepLab2

Change logs

Contacts (Maintainers)

Disclaimer

Citing DeepLab2

References

Owner

Google Research

Commonality in Natural Images Rescues GANs: Pretraining GANs with Generic and Privacy-free Synthetic Data - Official PyTorch Implementation (CVPR 2022)

A set of examples around hub for creating and processing datasets

Official implementation of deep-multi-trajectory-based single object tracking (IEEE T-CSVT 2021).

Autoregressive Predictive Coding: An unsupervised autoregressive model for speech representation learning

PyTorch code for our ECCV 2018 paper "Image Super-Resolution Using Very Deep Residual Channel Attention Networks"

Speech Emotion Recognition with Fusion of Acoustic- and Linguistic-Feature-Based Decisions

Official MegEngine implementation of CREStereo(CVPR 2022 Oral).

Multi-Modal Machine Learning toolkit based on PyTorch.

Structure-Preserving Deraining with Residue Channel Prior Guidance (ICCV2021)

Turning SymPy expressions into PyTorch modules.

automated systems to assist guarding corona Virus precautions for Closed Rooms (e.g. Halls, offices, etc..)

Physics-Aware Training (PAT) is a method to train real physical systems with backpropagation.

FCOSR: A Simple Anchor-free Rotated Detector for Aerial Object Detection

Code for the paper "Can Active Learning Preemptively Mitigate Fairness Issues?" presented at RAI 2021.

VISSL is FAIR's library of extensible, modular and scalable components for SOTA Self-Supervised Learning with images.

Feedback is important: response-aware feedback mechanism for background based conversation

A toolset of Python programs for signal modeling and indentification via sparse semilinear autoregressors.

Curriculum Domain Adaptation for Semantic Segmentation of Urban Scenes, ICCV 2017

π-GAN: Periodic Implicit Generative Adversarial Networks for 3D-Aware Image Synthesis

A simple API wrapper for Discord interactions.