A TensorFlow implementation of FCN-8s

Last update: Aug 08, 2022

Overview

FCN-8s implementation in TensorFlow

Overview
Examples and demo video
Dependencies
How to use it
Download pre-trained VGG-16

Overview

This is a TensorFlow implementation of the FCN-8s model architecture for semantic image segmentation introduced by Shelhamer et al. in the paper Fully Convolutional Networks for Semantic Segmentation.

This repository only contains the 'all-at-once' version of the FCN-8s model, which converges significantly faster than the version trained in stages. A convolutionalized VGG-16 model trained on ImageNet classification is provided and serves as the encoder of the FCN-8s. Sufficient documentation and a tutorial on how to train, evaluate and use the model for prediction are also provided. Some useful TensorBoard summaries can be recorded out of the box.

Examples and demo video

Below are some prediction examples of the model trained on the Cityscapes dataset for 13,000 steps at batch size 16, at which point the model achieves a mean IoU of 38.2% on the validation dataset. This is far from convergence of course, the purpose of these examples is just to demonstrate that the code works and the model learns. You can watch the model in action on the Cityscapes demo videos here.

Dependencies

Python 3.x
TensorFlow 1.x
Numpy
Scipy
OpenCV (for data augmentation)
tqdm

How to use it

fcn8s_tutorial.ipynb explains how to train and evaluate the model and how to make and visualize predictions.

Download pre-trained VGG-16

You can download the pre-trained, convolutionalized VGG-16 model here

A TensorFlow implementation of FCN-8s

Related tags

Overview

FCN-8s implementation in TensorFlow

Contents

Overview

Examples and demo video

Dependencies

How to use it

Download pre-trained VGG-16

Owner

Pierluigi Ferrari

This repository contains Prior-RObust Bayesian Optimization (PROBO) as introduced in our paper "Accounting for Gaussian Process Imprecision in Bayesian Optimization"

Structure-Preserving Deraining with Residue Channel Prior Guidance (ICCV2021)

Public implementation of "Learning from Suboptimal Demonstration via Self-Supervised Reward Regression" from CoRL'21

Predict bus arrival time using VertexAI and Nvidia's Jetson Nano

ToFFi - Toolbox for Frequency-based Fingerprinting of Brain Signals

Do Smart Glasses Dream of Sentimental Visions? Deep Emotionship Analysis for Eyewear Devices

Reinforcement Learning for Portfolio Management

Analyzes your GitHub Profile and presents you with a report on how likely you are to become the next MLH Fellow!

Node Dependent Local Smoothing for Scalable Graph Learning

A Deep Convolutional Encoder-Decoder Architecture for Image Segmentation

CSPML (crystal structure prediction with machine learning-based element substitution)

Deploy optimized transformer based models on Nvidia Triton server

Depression Asisstant GDSC Challenge Solution

Array Camera Ptychography

Serve TensorFlow ML models with TF-Serving and then create a Streamlit UI to use them

f-BRS: Rethinking Backpropagating Refinement for Interactive Segmentation

Source code for ZePHyR: Zero-shot Pose Hypothesis Rating @ ICRA 2021

Code for LIGA-Stereo Detector, ICCV'21

Optimizing synthesizer parameters using gradient approximation

PyTorch implementation of the NIPS-17 paper "Poincaré Embeddings for Learning Hierarchical Representations"