Clockwork Convnets for Video Semantic Segmentation

This is the reference implementation of arxiv:1608.03609:

Clockwork Convnets for Video Semantic Segmentation
Evan Shelhamer*, Kate Rakelly*, Judy Hoffman*, Trevor Darrell
arXiv:1605.06211

This project reproduces results from the arxiv and demonstrates how to execute staged fully convolutional networks (FCNs) on video in Caffe by controlling the net through the Python interface. In this way this these experiments are a proof-of-concept implementation of clockwork, and further development is needed to achieve peak efficiency (such as pre-fetching video data layers, threshold GPU layers, and a native Caffe library edition of the staged forward pass for pipelining).

For simple reference, refer to these (display only) editions of the experiments:

Cityscapes Clockwork
YouTube Frame Differencing
YouTube Clockwork
YouTube Pipelining
Synthetic PASCAL VOC Video
Dataset Walkthroughs for YouTube, NYUDv2, and Cityscapes

Contents

notebooks: interactive code and documentation that carries out the experiments (in jupyter/ipython format).
nets: the net specification of the various FCNs in this work, and the pre-trained weights (see installation instructions).
caffe: the Caffe framework, included as a git submodule pointing to a compatible version
datasets: input-output for PASCAL VOC, NYUDv2, YouTube-Objects, and Cityscapes
lib: helpers for executing networks, scoring metrics, and plotting

License

This project is licensed for open non-commercial distribution under the UC Regents license; see LICENSE. Its dependencies, such as Caffe, are subject to their own respective licenses.

Requirements & Installation

Caffe, Python, and Jupyter are necessary for all of the experiments. Any installation or general Caffe inquiries should be directed to the caffe-users mailing list.

Install Caffe. See the installation guide and try Caffe through Docker (recommended). Make sure to configure pycaffe, the Caffe Python interface, too.
Install Python, and then install our required packages listed in requirements.txt. For instance, for x in $(cat requirements.txt); do pip install $x; done should do.
Install Jupyter, the interface for viewing, executing, and altering the notebooks.
Configure your PYTHONPATH as indicated by the included .envrc so that this project dir and pycaffe are included.
Download the model weights for this project and place them in nets.

Now you can explore the notebooks by firing up Jupyter.

Clockwork Convnets for Video Semantic Segmentation

Related tags

Overview

Clockwork Convnets for Video Semantic Segmentation

License

Requirements & Installation

Owner

Evan Shelhamer

Data augmentation for NLP, accepted at EMNLP 2021 Findings

YOLOv3 in PyTorch > ONNX > CoreML > TFLite

(CVPR2021) Kaleido-BERT: Vision-Language Pre-training on Fashion Domain

Learning to Segment Instances in Videos with Spatial Propagation Network

Deep Residual Networks with 1K Layers

[CVPR'21] Projecting Your View Attentively: Monocular Road Scene Layout Estimation via Cross-view Transformation

Tensorflow port of a full NetVLAD network

Google AI Open Images - Object Detection Track: Open Solution

A gesture recognition system powered by OpenPose, k-nearest neighbours, and local outlier factor.

Advanced Deep Learning with TensorFlow 2 and Keras (Updated for 2nd Edition)

[ICCV'21] Official implementation for the paper Social NCE: Contrastive Learning of Socially-aware Motion Representations

VOS: Learning What You Don’t Know by Virtual Outlier Synthesis

An Implementation of Fully Convolutional Networks in Tensorflow.

Official Implementation of "LUNAR: Unifying Local Outlier Detection Methods via Graph Neural Networks"

[CVPR 2021] VirTex: Learning Visual Representations from Textual Annotations

Introducing neural networks to predict stock prices

MakeItTalk: Speaker-Aware Talking-Head Animation

Is RobustBench/AutoAttack a suitable Benchmark for Adversarial Robustness?

Hough Transform and Hough Line Transform Using OpenCV

Hyperopt for solving CIFAR-100 with a convolutional neural network (CNN) built with Keras and TensorFlow, GPU backend