Hierarchical Attentive Recurrent Tracking

Last update: Aug 07, 2021

Overview

Hierarchical Attentive Recurrent Tracking

This is an official Tensorflow implementation of single object tracking in videos by using hierarchical attentive recurrent neural networks, as presented in the following paper:

A. R. Kosiorek, A. Bewley, I. Posner, "Hierarchical Attentive Recurrent Tracking", NIPS 2017.

Author: Adam Kosiorek, Oxford Robotics Institue, University of Oxford
Email: adamk(at)robots.ox.ac.uk
Paper: https://arxiv.org/abs/1706.09262
Webpage: http://ori.ox.ac.uk/

Installation

Install Tensorflow v1.1 and the following dependencies (using pip install -r requirements.txt (preferred) or pip install [package]):

matplotlib==1.5.3
numpy==1.12.1
pandas==0.18.1
scipy==0.18.1

Demo

The notebook scripts/demo.ipynb contains a demo, which shows how to evaluate tracker on an arbitrary image sequence. By default, it runs on images located in imgs folder and uses a pretrained model. Before running the demo please download AlexNet weights first (described in the Training section).

Data

Download KITTI dataset from here. We need left color images and tracking labels.
Unpack data into a data folder; images should be in an image folder and labels should be in a label folder.
Resize all the images to (heigh=187, width=621) e.g. by using the scripts/resize_imgs.sh script.

Training

Download the AlexNet weights:
- Execute scripts/download_alexnet.sh or
- Download the weights from here and put the file in the checkpoints folder.

Run

 python scripts/train_hart_kitti.py --img_dir=path/to/image/folder --label_dir=/path/to/label/folder

The training script will save model checkpoints in the checkpoints folder and report train and test scores every couple of epochs. You can run tensorboard in the checkpoints folder to visualise training progress. Training should converge in about 400k iterations, which should take about 3 days. It might take a couple of hours between logging messages, so don't worry.

Evaluation on KITTI dataset

The scripts/eval_kitti.ipynb notebook contains the code necessary to prepare (IoU, timesteps) curves for train and validation set of KITTI. Before running the evaluation:

Download AlexNet weights (described in the Training section).
Update image and label folder paths in the notebook.

Citation

If you find this repo useful in your research, please consider citing:

@inproceedings{Kosiorek2017hierarchical,
   title = {Hierarchical Attentive Recurrent Tracking},
   author = {Kosiorek, Adam R and Bewley, Alex and Posner, Ingmar},
   booktitle = {Neural Information Processing Systems},
   url = {http://www.robots.ox.ac.uk/~mobile/Papers/2017NIPS_AdamKosiorek.pdf},
   pdf = {http://www.robots.ox.ac.uk/~mobile/Papers/2017NIPS_AdamKosiorek.pdf},
   year = {2017},
   month = {December}
}

License

This program is free software; you can redistribute it and/or modify it under the terms of the GNU General Public License as published by the Free Software Foundation; either version 3 of the License, or (at your option) any later version.

This program is distributed in the hope that it will be useful, but WITHOUT ANY WARRANTY; without even the implied warranty of MERCHANTABILITY or FITNESS FOR A PARTICULAR PURPOSE. See the GNU General Public License for more details.

You should have received a copy of the GNU General Public License along with this program. If not, see http://www.gnu.org/licenses/.

Release Notes

Version 1.0

Original version from the paper. It contains the KITTI tracking experiment.

Hierarchical Attentive Recurrent Tracking

Related tags

Overview

Hierarchical Attentive Recurrent Tracking

Installation

Demo

Data

Training

Evaluation on KITTI dataset

Citation

License

Release Notes

Owner

Adam Kosiorek

An Ensemble of CNN (Python 3.5.1 Tensorflow 1.3 numpy 1.13)

A stock generator that assess a list of stocks and returns the best stocks for investing and money allocations based on users choices of volatility, duration and number of stocks

PyTorch implementation of PNASNet-5 on ImageNet

Jetson Nano-based smart camera system that measures crowd face mask usage in real-time.

WiFi-based Multi-task Sensing

coldcuts is an R package to automatically generate and plot segmentation drawings in R

SelfAugment extends MoCo to include automatic unsupervised augmentation selection.

Solution to the first stage Quiz of Hamoye internship: Introduction to Python for Machine Learning

(JMLR' 19) A Python Toolbox for Scalable Outlier Detection (Anomaly Detection)

Code for our ACL 2021 paper - ConSERT: A Contrastive Framework for Self-Supervised Sentence Representation Transfer

AoT is a system for automatically generating off-target test harness by using build information.

Official PyTorch implementation of "Proxy Synthesis: Learning with Synthetic Classes for Deep Metric Learning" (AAAI 2021)

Toolkit for collecting and applying prompts

Easily benchmark PyTorch model FLOPs, latency, throughput, max allocated memory and energy consumption

This is the repository for the paper "Have I done enough planning or should I plan more?"

Additional environments compatible with OpenAI gym

iris - Open Source Photos Platform Powered by PyTorch

Bayesian-Torch is a library of neural network layers and utilities extending the core of PyTorch to enable the user to perform stochastic variational inference in Bayesian deep neural networks

Boosting Monocular Depth Estimation Models to High-Resolution via Content-Adaptive Multi-Resolution Merging

This is a project based on ConvNets used to identify whether a road is clean or dirty. We have used MobileNet as our base architecture and the weights are based on imagenet.