Code for CVPR 2022 paper "SoftGroup for Instance Segmentation on 3D Point Clouds"

Last update: Dec 27, 2022

Related tags

Overview

SoftGroup

We provide code for reproducing results of the paper SoftGroup for 3D Instance Segmentation on Point Clouds (CVPR 2022)

Author: Thang Vu, Kookhoi Kim, Tung M. Luu, Xuan Thanh Nguyen, and Chang D. Yoo.

Introduction

Existing state-of-the-art 3D instance segmentation methods perform semantic segmentation followed by grouping. The hard predictions are made when performing semantic segmentation such that each point is associated with a single class. However, the errors stemming from hard decision propagate into grouping that results in (1) low overlaps between the predicted instance with the ground truth and (2) substantial false positives. To address the aforementioned problems, this paper proposes a 3D instance segmentation method referred to as SoftGroup by performing bottom-up soft grouping followed by top-down refinement. SoftGroup allows each point to be associated with multiple classes to mitigate the problems stemming from semantic prediction errors and suppresses false positive instances by learning to categorize them as background. Experimental results on different datasets and multiple evaluation metrics demonstrate the efficacy of SoftGroup. Its performance surpasses the strongest prior method by a significant margin of +6.2% on the ScanNet v2 hidden test set and +6.8% on S3DIS Area 5 of AP_50.

Feature

State of the art performance on the ScanNet benchmark and S3DIS dataset (3/Mar/2022).
High speed of 345 ms per scan on ScanNet dataset, which is comparable with the existing fastest methods (HAIS).
Reproducibility code for both ScanNet and S3DIS datasets.

Installation

Please refer to installation guide.

Data Preparation

Please refer to data preparation for preparing the S3DIS and ScanNet v2 dataset.

Pretrained models

Dataset	AP	AP_50	AP_25	Download
S3DIS	51.4	66.5	75.4	model
ScanNet v2	46.0	67.6	78.9	model

Training

We use the checkpoint of HAIS as pretrained backbone. Download the pretrained HAIS model at here at put it in SoftGroup/ directory.

Training S3DIS dataset

First, finetune the pretrained HAIS point-wise prediction network (backbone) on S3DIS.

python train.py --config config/softgroup_fold5_backbone_s3dis.yaml

Then, train model from frozen backbone.

python train.py --config config/softgroup_fold5_default_s3dis.yaml

Training ScanNet V2 dataset

Training on ScanNet doesnot require finetuning the backbone. Just freeze pretrained backbone and train the model.

python train.py --config config/softgroup_default_scannet.yaml

Inference

Testing for S3DIS dataset.

CUDA_VISIBLE_DEVICES=0 python test_s3dis.py --config config/softgroup_fold5_phase2_s3dis.yaml --pretrain $PATH_TO_PRETRAIN_MODEL$

Testing for ScanNet V2 dataset.

CUDA_VISIBLE_DEVICES=0 python test.py --config config/softgroup_default_scannet.yaml --pretrain $PATH_TO_PRETRAIN_MODEL$

Visualization

We provide visualization tools based on Open3D (tested on Open3D 0.8.0).

pip install open3D==0.8.0
python visualize_open3d.py --data_path {} --prediction_path {} --data_split {} --room_name {} --task {}

Please refer to visualize_open3d.py for more details.

Citation

If you find our work helpful for your research. Please consider citing our paper.

@inproceedings{vu2022softgroup,
  title={SoftGroup for 3D Instance Segmentation on 3D Point Clouds},
  author={Vu, Thang and Kim, Kookhoi and Luu, Tung M. and Nguyen, Xuan Thanh and Yoo, Chang D.},
  booktitle={CVPR},
  year={2022}
}

Code for CVPR 2022 paper "SoftGroup for Instance Segmentation on 3D Point Clouds"

Related tags

Overview

SoftGroup

Introduction

Feature

Installation

Data Preparation

Pretrained models

Training

Training S3DIS dataset

Training ScanNet V2 dataset

Inference

Visualization

Citation

Owner

Thang Vu

Demo processor to illustrate OCR-D Python API

Turn images of tables into CSV data. Detect tables from images and run OCR on the cells.

Python Computer Vision Aim Bot for Roblox's Phantom Forces

Hiiii this is the Spanish for Linux and win 10 and in the near future the english version of PortScan my new tool on which you can see what ports are Open only with the IP adress.

Code for the head detector (HeadHunter) proposed in our CVPR 2021 paper Tracking Pedestrian Heads in Dense Crowd.

Python Computer Vision application that allows users to draw/erase on the screen using their webcam.

Deep learning based page layout analysis

Contextual speed detection for python

Genalog is an open source, cross-platform python package allowing generation of synthetic document images with custom degradations and text alignment capabilities.

SRA's seminar on Introduction to Computer Vision Fundamentals

Bu uygulamada Python ve Opencv kullanarak bilgisayar kamerasından yüz tespiti yapıyoruz.

This is a implementation of CRAFT OCR method

A version of nrsc5-gui that merges the interface developed by cmnybo with the architecture developed by zefie in order to start a new baseline that is not heavily dependent upon Python processing.

Virtualdragdrop - Virtual Drag and Drop Using OpenCV and Arduino

An interactive interface for using OpenCV's GrabCut algorithm for image segmentation.

Code for the paper STN-OCR: A single Neural Network for Text Detection and Text Recognition

Perspective recovery of text using transformed ellipses

Indonesian ID Card OCR using tesseract OCR

Awesome Spectral Indices in Python.

かの有名なあの東方二次創作ソング、「bad apple!」のMVをPythonでやってみたって話