CTRL-C: Camera calibration TRansformer with Line-Classification

Last update: Nov 14, 2022

Related tags

Overview

CTRL-C: Camera calibration TRansformer with Line-Classification

This repository contains the official code and pretrained models for CTRL-C (Camera calibration TRansformer with Line-Classification). Jinwoo Lee, Hyunsung Go, Hyunjoon Lee, Sunghyun Cho, Minhyuk Sung and Junho Kim. ICCV 2021.

Single image camera calibration is the task of estimating the camera parameters from a single input image, such as the vanishing points, focal length, and horizon line. In this work, we propose Camera calibration TRansformer with Line-Classification (CTRL-C), an end-to-end neural network-based approach to single image camera calibration, which directly estimates the camera parameters from an image and a set of line segments. Our network adopts the transformer architecture to capture the global structure of an image with multi-modal inputs in an end-to-end manner. We also propose an auxiliary task of line classification to train the network to extract the global geometric information from lines effectively. Our experiments demonstrate that CTRL-C outperforms the previous state-of-the-art methods on the Google Street View and SUN360 benchmark datasets.

Results & Checkpoints

Dataset	Up Dir (◦)	Pitch (◦)	Roll (◦)	FoV (◦)	AUC (%)	URL
Google Street View	1.80	1.58	0.66	3.59	87.29	gdrive
SUN360	1.91	1.50	0.96	3.80	85.45	gdrive

Preparation

Clone this repository

Setup environments

conda create -n ctrlc python
conda activate ctrlc
conda install -c pytorch torchvision

pip install -r requrements.txt

Training Datasets

Google Street View dataset
SUN360 dataset
- You need to preprocess the dataset

Training

Single GPU

python main.py --config-file 'config-files/ctrl-c.yaml' --opts OUTPUT_DIR 'logs'

Multi GPU

python -m torch.distributed.launch --nproc_per_node=4 --use_env main.py --config-file 'config-files/ctrl-c.yaml' --opts OUTPUT_DIR 'logs'

Evaluation

python test.py --dataset 'GoogleStreetView' --opts OUTPUT_DIR 'outputs'

Citation

If you use this code for your research, please cite our paper:

@InProceedings{Lee:2021:ICCV,
    Title     = {{CTRL-C: Camera calibration TRansformer with Line-Classification}},
    Author    = {Jinwoo Lee and Hyunsung Go and Hyunjoon Lee and Sunghyun Cho and Minhyuk Sung and Junho Kim},    
    Booktitle = {Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV)},
    Year      = {2021},
}

License

CTRL-C is released under the Apache 2.0 license. Please see the LICENSE file for more information.

Acknowledgments

This code is based on the implementations of DETR: End-to-End Object Detection with Transformers.

CTRL-C: Camera calibration TRansformer with Line-Classification

Related tags

Overview

CTRL-C: Camera calibration TRansformer with Line-Classification

Results & Checkpoints

Preparation

Training Datasets

Training

Evaluation

Citation

License

Acknowledgments

Owner

Train CNNs for the fruits360 data set in NTOU CS「Machine Vision」class.

Selene is a Python library and command line interface for training deep neural networks from biological sequence data such as genomes.

Explaining neural decisions contrastively to alternative decisions.

Open source repository for the code accompanying the paper 'PatchNets: Patch-Based Generalizable Deep Implicit 3D Shape Representations'.

Official PyTorch Implementation of HELP: Hardware-adaptive Efficient Latency Prediction for NAS via Meta-Learning (NeurIPS 2021 Spotlight)

An alarm clock coded in Python 3 with Tkinter

Fashion Entity Classification

[ICLR 2021] "CPT: Efficient Deep Neural Network Training via Cyclic Precision" by Yonggan Fu, Han Guo, Meng Li, Xin Yang, Yining Ding, Vikas Chandra, Yingyan Lin

Source code for our paper "Learning to Break Deep Perceptual Hashing: The Use Case NeuralHash"

A PyTorch implementation of "SelfGNN: Self-supervised Graph Neural Networks without explicit negative sampling"

[CVPR 2020] Interpreting the Latent Space of GANs for Semantic Face Editing

You Only Sample (Almost) Once: Linear Cost Self-Attention Via Bernoulli Sampling

Improving adversarial robustness by a coupling rejection strategy

Rayvens makes it possible for data scientists to access hundreds of data services within Ray with little effort.

CenterNet:Objects as Points目标检测模型在Pytorch当中的实现

A Pytorch Implementation of ClariNet

Civsim is a basic civilisation simulation and modelling system built in Python 3.8.

A universal framework for learning timestamp-level representations of time series

Python scripts to detect faces in Python with the BlazeFace Tensorflow Lite models

IDA file loader for UF2, created for the DEFCON 29 hardware badge