CTRL-C: Camera calibration TRansformer with Line-Classification

Last update: Nov 14, 2022

Related tags

Overview

CTRL-C: Camera calibration TRansformer with Line-Classification

This repository contains the official code and pretrained models for CTRL-C (Camera calibration TRansformer with Line-Classification). Jinwoo Lee, Hyunsung Go, Hyunjoon Lee, Sunghyun Cho, Minhyuk Sung and Junho Kim. ICCV 2021.

Single image camera calibration is the task of estimating the camera parameters from a single input image, such as the vanishing points, focal length, and horizon line. In this work, we propose Camera calibration TRansformer with Line-Classification (CTRL-C), an end-to-end neural network-based approach to single image camera calibration, which directly estimates the camera parameters from an image and a set of line segments. Our network adopts the transformer architecture to capture the global structure of an image with multi-modal inputs in an end-to-end manner. We also propose an auxiliary task of line classification to train the network to extract the global geometric information from lines effectively. Our experiments demonstrate that CTRL-C outperforms the previous state-of-the-art methods on the Google Street View and SUN360 benchmark datasets.

Results & Checkpoints

Dataset	Up Dir (◦)	Pitch (◦)	Roll (◦)	FoV (◦)	AUC (%)	URL
Google Street View	1.80	1.58	0.66	3.59	87.29	gdrive
SUN360	1.91	1.50	0.96	3.80	85.45	gdrive

Preparation

Clone this repository

Setup environments

conda create -n ctrlc python
conda activate ctrlc
conda install -c pytorch torchvision

pip install -r requrements.txt

Training Datasets

Google Street View dataset
SUN360 dataset
- You need to preprocess the dataset

Training

Single GPU

python main.py --config-file 'config-files/ctrl-c.yaml' --opts OUTPUT_DIR 'logs'

Multi GPU

python -m torch.distributed.launch --nproc_per_node=4 --use_env main.py --config-file 'config-files/ctrl-c.yaml' --opts OUTPUT_DIR 'logs'

Evaluation

python test.py --dataset 'GoogleStreetView' --opts OUTPUT_DIR 'outputs'

Citation

If you use this code for your research, please cite our paper:

@InProceedings{Lee:2021:ICCV,
    Title     = {{CTRL-C: Camera calibration TRansformer with Line-Classification}},
    Author    = {Jinwoo Lee and Hyunsung Go and Hyunjoon Lee and Sunghyun Cho and Minhyuk Sung and Junho Kim},    
    Booktitle = {Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV)},
    Year      = {2021},
}

License

CTRL-C is released under the Apache 2.0 license. Please see the LICENSE file for more information.

Acknowledgments

This code is based on the implementations of DETR: End-to-End Object Detection with Transformers.

CTRL-C: Camera calibration TRansformer with Line-Classification

Related tags

Overview

CTRL-C: Camera calibration TRansformer with Line-Classification

Results & Checkpoints

Preparation

Training Datasets

Training

Evaluation

Citation

License

Acknowledgments

Owner

A library for researching neural networks compression and acceleration methods.

A style-based Quantum Generative Adversarial Network

TensorFlow CNN for fast style transfer

[ECCV2020] Content-Consistent Matching for Domain Adaptive Semantic Segmentation

Official Pytorch implementation of "Learning to Estimate Robust 3D Human Mesh from In-the-Wild Crowded Scenes", CVPR 2022

This dlib-based facial login system

Python scripts for performing object detection with the 1000 labels of the ImageNet dataset in ONNX.

Train neural network for semantic segmentation (deep lab V3) with pytorch in less then 50 lines of code

Code needed to reproduce the examples found in "The Temporal Robustness of Stochastic Signals"

Bayesian Deep Learning and Deep Reinforcement Learning for Object Shape Error Response and Correction of Manufacturing Systems

text_recognition_toolbox: The reimplementation of a series of classical scene text recognition papers with Pytorch in a uniform way.

Neural HMMs are all you need (for high-quality attention-free TTS)

A tool to estimate time varying instantaneous reproduction number during epidemics

An open software package to develop BCI based brain and cognitive computing technology for recognizing user's intention using deep learning

Code and real data for the paper "Counterfactual Temporal Point Processes", available at arXiv.

A novel pipeline framework for multi-hop complex KGQA task. About the paper title: Improving Multi-hop Embedded Knowledge Graph Question Answering by Introducing Relational Chain Reasoning

Official implementation of the paper "Lightweight Deep CNN for Natural Image Matting via Similarity Preserving Knowledge Distillation"

Notebooks for my "Deep Learning with TensorFlow 2 and Keras" course

Source code for Adaptively Calibrated Critic Estimates for Deep Reinforcement Learning

Uncertainty-aware Semantic Segmentation of LiDAR Point Clouds for Autonomous Driving