Semantic segmentation task for ADE20k & cityscapse dataset, based on several models.

Last update: Oct 13, 2022

Overview

semantic-segmentation-tensorflow

This is a Tensorflow implementation of semantic segmentation models on MIT ADE20K scene parsing dataset and Cityscapes dataset We re-produce the inference phase of several models, including PSPNet, FCN, and ICNet by transforming the released pre-trained weights into tensorflow format, and apply on handcraft models. Also, we refer to ENet from freg856 github. Still working on task integrated.

Models

PSPNet
FCN
ENet
ICNet

...to be continue

Install

Get corresponding transformed pre-trained weights, and put into model directory:

FCN	PSPNet	ICNet
Google drive	Google drive	Google drive

Inference

Run following command:

python inference.py --img-path /Path/To/Image --dataset Model_Type

Arg list

--model - choose from "icnet"/"pspnet"/"fcn"/"enet"

Import module in your code:

from model import FCN8s, PSPNet50, ICNet, ENet

model = PSPNet50() # or another model

model.read_input(img_path)  # read image data from path

sess = tf.Session(config=config)
init = tf.global_variables_initializer()
sess.run(init)

model.load(model_path, sess)  # load pretrained model
preds = model.forward(sess) # Get prediction

Results

ade20k

Input Image	PSPNet	FCN

cityscapes

Input Image	ICNet	ENet

Citation

@inproceedings{zhao2017pspnet,
  author = {Hengshuang Zhao and
            Jianping Shi and
            Xiaojuan Qi and
            Xiaogang Wang and
            Jiaya Jia},
  title = {Pyramid Scene Parsing Network},
  booktitle = {Proceedings of IEEE Conference on Computer Vision and Pattern Recognition (CVPR)},
  year = {2017}
}

Scene Parsing through ADE20K Dataset. B. Zhou, H. Zhao, X. Puig, S. Fidler, A. Barriuso and A. Torralba. Computer Vision and Pattern Recognition (CVPR), 2017. (http://people.csail.mit.edu/bzhou/publication/scene-parse-camera-ready.pdf)

@inproceedings{zhou2017scene,
    title={Scene Parsing through ADE20K Dataset},
    author={Zhou, Bolei and Zhao, Hang and Puig, Xavier and Fidler, Sanja and Barriuso, Adela and Torralba, Antonio},
    booktitle={Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition},
    year={2017}
}

Semantic Understanding of Scenes through ADE20K Dataset. B. Zhou, H. Zhao, X. Puig, S. Fidler, A. Barriuso and A. Torralba. arXiv:1608.05442. (https://arxiv.org/pdf/1608.05442.pdf)

@article{zhou2016semantic,
  title={Semantic understanding of scenes through the ade20k dataset},
  author={Zhou, Bolei and Zhao, Hang and Puig, Xavier and Fidler, Sanja and Barriuso, Adela and Torralba, Antonio},
  journal={arXiv preprint arXiv:1608.05442},
  year={2016}
}

Semantic segmentation task for ADE20k & cityscapse dataset, based on several models.

Related tags

Overview

semantic-segmentation-tensorflow

Models

...to be continue

Install

Inference

Arg list

Import module in your code:

Results

ade20k

cityscapes

Citation

Owner

HsuanKung Yang

Deep Learning tutorials in jupyter notebooks.

Official implementation for paper Render In-between: Motion Guided Video Synthesis for Action Interpolation

Code for our SIGCOMM'21 paper "Network Planning with Deep Reinforcement Learning".

The official implementation of NeMo: Neural Mesh Models of Contrastive Features for Robust 3D Pose Estimation [ICLR-2021]. https://arxiv.org/pdf/2101.12378.pdf

TensorFlow implementation of "TokenLearner: What Can 8 Learned Tokens Do for Images and Videos?"

Reproduce partial features of DeePMD-kit using PyTorch.

A PyTorch-based Semi-Supervised Learning (SSL) Codebase for Pixel-wise (Pixel) Vision Tasks

Code for our ICASSP 2021 paper: SA-Net: Shuffle Attention for Deep Convolutional Neural Networks

Local Attention - Flax module for Jax

Expand human face editing via Global Direction of StyleCLIP, especially to maintain similarity during editing.

HomeAssitant custom integration for dyson

Tweesent-back - Tweesent backend uses fastAPI as the web framework

Code Release for Learning to Adapt to Evolving Domains

Implementation of the Chamfer Distance as a module for pyTorch

Using deep actor-critic model to learn best strategies in pair trading

Harmonious Textual Layout Generation over Natural Images via Deep Aesthetics Learning

Monocular Depth Estimation - Weighted-average prediction from multiple pre-trained depth estimation models

This is the source code for the experiments related to the paper Unsupervised Audio Source Separation Using Differentiable Parametric Source Models

PyElastica is the Python implementation of Elastica, an open-source software for the simulation of assemblies of slender, one-dimensional structures using Cosserat Rod theory.

Code implementation for the paper 'Conditional Gaussian PAC-Bayes'.