Implementation of SegNet: A Deep Convolutional Encoder-Decoder Architecture for Semantic Pixel-Wise Labelling

Last update: Jan 02, 2023

Related tags

Overview

Caffe SegNet

This is a modified version of Caffe which supports the SegNet architecture

As described in SegNet: A Deep Convolutional Encoder-Decoder Architecture for Image Segmentation Vijay Badrinarayanan, Alex Kendall and Roberto Cipolla, PAMI 2017 [http://arxiv.org/abs/1511.00561]

Updated Version:

This version supports cudnn v2 acceleration. @TimoSaemann has a branch supporting a more recent version of Caffe (Dec 2016) with cudnn v5.1: https://github.com/TimoSaemann/caffe-segnet-cudnn5

Getting Started with Example Model and Webcam Demo

If you would just like to try out a pretrained example model, then you can find the model used in the SegNet webdemo and a script to run a live webcam demo here: https://github.com/alexgkendall/SegNet-Tutorial

For a more detailed introduction to this software please see the tutorial here: http://mi.eng.cam.ac.uk/projects/segnet/tutorial.html

Dataset

Prepare a text file of space-separated paths to images (jpegs or pngs) and corresponding label images alternatively e.g. /path/to/im1.png /another/path/to/lab1.png /path/to/im2.png /path/lab2.png ...

Label images must be single channel, with each value from 0 being a separate class. The example net uses an image size of 360 by 480.

Net specification

Example net specification and solver prototext files are given in examples/segnet. To train a model, alter the data path in the data layers in net.prototxt to be your dataset.txt file (as described above).

In the last convolution layer, change num_output to be the number of classes in your dataset.

Training

In solver.prototxt set a path for snapshot_prefix. Then in a terminal run ./build/tools/caffe train -solver ./examples/segnet/solver.prototxt

Publications

If you use this software in your research, please cite our publications:

http://arxiv.org/abs/1511.02680 Alex Kendall, Vijay Badrinarayanan and Roberto Cipolla "Bayesian SegNet: Model Uncertainty in Deep Convolutional Encoder-Decoder Architectures for Scene Understanding." arXiv preprint arXiv:1511.02680, 2015.

http://arxiv.org/abs/1511.00561 Vijay Badrinarayanan, Alex Kendall and Roberto Cipolla "SegNet: A Deep Convolutional Encoder-Decoder Architecture for Image Segmentation." PAMI, 2017.

License

This extension to the Caffe library is released under a creative commons license which allows for personal and research use only. For a commercial license please contact the authors. You can view a license summary here: http://creativecommons.org/licenses/by-nc/4.0/

Implementation of SegNet: A Deep Convolutional Encoder-Decoder Architecture for Semantic Pixel-Wise Labelling

Related tags

Overview

Caffe SegNet

Updated Version:

Getting Started with Example Model and Webcam Demo

Dataset

Net specification

Training

Publications

License

Owner

Alex Kendall

End-to-end image segmentation kit based on PaddlePaddle.

INSPIRED: A Transparent Dialogue Dataset for Interactive Semantic Parsing

Create time-series datacubes for supervised machine learning with ICEYE SAR images.

Source code of all the projects of Udacity Self-Driving Car Engineer Nanodegree.

MAT: Mask-Aware Transformer for Large Hole Image Inpainting

Official Implementation of CoSMo: Content-Style Modulation for Image Retrieval with Text Feedback

Causal Influence Detection for Improving Efficiency in Reinforcement Learning

Official implement of Paper：A deeply supervised image fusion network for change detection in high resolution bi-temporal remote sening images

Code for sound field predictions in domains with impedance boundaries. Used for generating results from the paper

Deep Learning and Logical Reasoning from Data and Knowledge

Visualizing Yolov5's layers using GradCam

Current state of supervised and unsupervised depth completion methods

MazeRL is an application oriented Deep Reinforcement Learning (RL) framework

Flax is a neural network ecosystem for JAX that is designed for flexibility.

Simple Tensorflow implementation of Toward Spatially Unbiased Generative Models (ICCV 2021)

A PyTorch implementation for V-Net: Fully Convolutional Neural Networks for Volumetric Medical Image Segmentation

Asterisk is a framework to generate high-quality training datasets at scale

The repository includes the code for training cell counting applications. (Keras + Tensorflow)

This is a GUI interface which can process forest fire detection, smoke detection and fire segmentation

Implementing yolov4 target detection and tracking based on nao robot