Clockwork Convnets for Video Semantic Segmentation

This is the reference implementation of arxiv:1608.03609:

Clockwork Convnets for Video Semantic Segmentation
Evan Shelhamer*, Kate Rakelly*, Judy Hoffman*, Trevor Darrell
arXiv:1605.06211

This project reproduces results from the arxiv and demonstrates how to execute staged fully convolutional networks (FCNs) on video in Caffe by controlling the net through the Python interface. In this way this these experiments are a proof-of-concept implementation of clockwork, and further development is needed to achieve peak efficiency (such as pre-fetching video data layers, threshold GPU layers, and a native Caffe library edition of the staged forward pass for pipelining).

For simple reference, refer to these (display only) editions of the experiments:

Cityscapes Clockwork
YouTube Frame Differencing
YouTube Clockwork
YouTube Pipelining
Synthetic PASCAL VOC Video
Dataset Walkthroughs for YouTube, NYUDv2, and Cityscapes

Contents

notebooks: interactive code and documentation that carries out the experiments (in jupyter/ipython format).
nets: the net specification of the various FCNs in this work, and the pre-trained weights (see installation instructions).
caffe: the Caffe framework, included as a git submodule pointing to a compatible version
datasets: input-output for PASCAL VOC, NYUDv2, YouTube-Objects, and Cityscapes
lib: helpers for executing networks, scoring metrics, and plotting

License

This project is licensed for open non-commercial distribution under the UC Regents license; see LICENSE. Its dependencies, such as Caffe, are subject to their own respective licenses.

Requirements & Installation

Caffe, Python, and Jupyter are necessary for all of the experiments. Any installation or general Caffe inquiries should be directed to the caffe-users mailing list.

Install Caffe. See the installation guide and try Caffe through Docker (recommended). Make sure to configure pycaffe, the Caffe Python interface, too.
Install Python, and then install our required packages listed in requirements.txt. For instance, for x in $(cat requirements.txt); do pip install $x; done should do.
Install Jupyter, the interface for viewing, executing, and altering the notebooks.
Configure your PYTHONPATH as indicated by the included .envrc so that this project dir and pycaffe are included.
Download the model weights for this project and place them in nets.

Now you can explore the notebooks by firing up Jupyter.

Clockwork Convnets for Video Semantic Segmentation

Related tags

Overview

Clockwork Convnets for Video Semantic Segmentation

License

Requirements & Installation

Owner

Evan Shelhamer

The datasets and code of ACL 2021 paper "Aspect-Category-Opinion-Sentiment Quadruple Extraction with Implicit Aspects and Opinions".

Official PyTorch implementation of the paper "Deep Constrained Least Squares for Blind Image Super-Resolution", CVPR 2022.

Pytorch implementation code for [Neural Architecture Search for Spiking Neural Networks]

Vector Quantized Diffusion Model for Text-to-Image Synthesis

Python scripts for performing object detection with the 1000 labels of the ImageNet dataset in ONNX.

Get the partition that a file belongs and the percentage of space that consumes

A modular, open and non-proprietary toolkit for core robotic functionalities by harnessing deep learning

Scalable Multi-Agent Reinforcement Learning

VolumeGAN - 3D-aware Image Synthesis via Learning Structural and Textural Representations

Simple-Image-Classification - Simple Image Classification Code (PyTorch)

Complete U-net Implementation with keras

[IJCAI-2021] A benchmark of data-free knowledge distillation from paper "Contrastive Model Inversion for Data-Free Knowledge Distillation"

Learning Visual Words for Weakly-Supervised Semantic Segmentation

Decentralized Reinforcment Learning: Global Decision-Making via Local Economic Transactions (ICML 2020)

Data and analysis code for an MS on SK VOC genomes phenotyping/neutralisation assays

[ICCV 2021 (oral)] Planar Surface Reconstruction from Sparse Views

MicroNet: Improving Image Recognition with Extremely Low FLOPs (ICCV 2021)

Privacy-Preserving Machine Learning (PPML) Tutorial Presented at PyConDE 2022

Implementation of the paper "Self-Promoted Prototype Refinement for Few-Shot Class-Incremental Learning"

TrTr: Visual Tracking with Transformer