Clockwork Convnets for Video Semantic Segmentation

This is the reference implementation of arxiv:1608.03609:

Clockwork Convnets for Video Semantic Segmentation
Evan Shelhamer*, Kate Rakelly*, Judy Hoffman*, Trevor Darrell
arXiv:1605.06211

This project reproduces results from the arxiv and demonstrates how to execute staged fully convolutional networks (FCNs) on video in Caffe by controlling the net through the Python interface. In this way this these experiments are a proof-of-concept implementation of clockwork, and further development is needed to achieve peak efficiency (such as pre-fetching video data layers, threshold GPU layers, and a native Caffe library edition of the staged forward pass for pipelining).

For simple reference, refer to these (display only) editions of the experiments:

Cityscapes Clockwork
YouTube Frame Differencing
YouTube Clockwork
YouTube Pipelining
Synthetic PASCAL VOC Video
Dataset Walkthroughs for YouTube, NYUDv2, and Cityscapes

Contents

notebooks: interactive code and documentation that carries out the experiments (in jupyter/ipython format).
nets: the net specification of the various FCNs in this work, and the pre-trained weights (see installation instructions).
caffe: the Caffe framework, included as a git submodule pointing to a compatible version
datasets: input-output for PASCAL VOC, NYUDv2, YouTube-Objects, and Cityscapes
lib: helpers for executing networks, scoring metrics, and plotting

License

This project is licensed for open non-commercial distribution under the UC Regents license; see LICENSE. Its dependencies, such as Caffe, are subject to their own respective licenses.

Requirements & Installation

Caffe, Python, and Jupyter are necessary for all of the experiments. Any installation or general Caffe inquiries should be directed to the caffe-users mailing list.

Install Caffe. See the installation guide and try Caffe through Docker (recommended). Make sure to configure pycaffe, the Caffe Python interface, too.
Install Python, and then install our required packages listed in requirements.txt. For instance, for x in $(cat requirements.txt); do pip install $x; done should do.
Install Jupyter, the interface for viewing, executing, and altering the notebooks.
Configure your PYTHONPATH as indicated by the included .envrc so that this project dir and pycaffe are included.
Download the model weights for this project and place them in nets.

Now you can explore the notebooks by firing up Jupyter.

Clockwork Convnets for Video Semantic Segmentation

Related tags

Overview

Clockwork Convnets for Video Semantic Segmentation

License

Requirements & Installation

Owner

Evan Shelhamer

A SAT-based sudoku solver

To SMOTE, or not to SMOTE?

Tensorflow port of a full NetVLAD network

Dilated RNNs in pytorch

OpenMMLab's Next Generation Video Understanding Toolbox and Benchmark

This is project is the implementation of the DeepShift: Towards Multiplication-Less Neural Networks paper

Repository for Traffic Accident Benchmark for Causality Recognition (ECCV 2020)

Simple and Effective Few-Shot Named Entity Recognition with Structured Nearest Neighbor Learning

An onlinel learning to rank python codebase.

Language Used: Python . Made in Jupyter(Anaconda) notebook.

ICLR 2021, Fair Mixup: Fairness via Interpolation

这是一个unet-pytorch的源码，可以训练自己的模型

A curated list of references for MLOps

Official PyTorch implementation of the paper: DeepSIM: Image Shape Manipulation from a Single Augmented Training Sample

Calculates carbon footprint based on fuel mix and discharge profile at the utility selected. Can create graphs and tabular output for fuel mix based on input file of series of power drawn over a period of time.

A simple software for capturing human body movements using the Kinect camera.

VLG-Net: Video-Language Graph Matching Networks for Video Grounding

Python scripts for performing road segemtnation and car detection using the HybridNets multitask model in ONNX.

Cross-Document Coreference Resolution

Learning To Have An Ear For Face Super-Resolution