Global-Local Path Networks for Monocular Depth Estimation with Vertical CutDepth [Paper]

Last update: Jan 01, 2023

Related tags

Deep Learning GLPDepth

Overview

Global-Local Path Networks for Monocular Depth Estimation with Vertical CutDepth [Paper]

Downloads

[Downloads] Trained ckpt files for NYU Depth V2 and KITTI
[Downloads] Predicted depth maps png files for NYU Depth V2 and KITTI Eigen split test set

Requirements

Tested on

python==3.7.7
torch==1.6.0
h5py==3.6.0
scipy==1.7.3
opencv-python==4.5.5
mmcv==1.4.3
timm=0.5.4
albumentations=1.1.0
tensorboardX==2.4.1

You can install above package with

$ pip install -r requirements.txt

Inference and Evaluate

Dataset

NYU Depth V2

$ cd ./datasets
$ wget http://horatio.cs.nyu.edu/mit/silberman/nyu_depth_v2/nyu_depth_v2_labeled.mat
$ python ../code/utils/extract_official_train_test_set_from_mat.py nyu_depth_v2_labeled.mat splits.mat ./nyu_depth_v2/official_splits/

KITTI

Download annotated depth maps data set (14GB) from [link] into ./datasets/kitti/data_depth_annotated

$ cd ./datasets/kitti/data_depth_annotated/
$ unzip data_depth_annotated.zip

With above two instrtuctions, you can perform eval_with_pngs.py/test.py for NYU Depth V2 and eval_with_pngs for KITTI.

To fully perform experiments, please follow [BTS] repository to obtain full dataset for NYU Depth V2 and KITTI datasets.

Your dataset directory should be

root
- nyu_depth_v2
  - bathroom_0001
  - bathroom_0002
  - ...
  - official_splits
- kitti
  - data_depth_annotated
  - raw_data
  - val_selection_cropped

Evaluation

Evaluate with png images

for NYU Depth V2

$ python ./code/eval_with_pngs.py --dataset nyudepthv2 --pred_path ./best_nyu_preds/ --gt_path ./datasets/nyu_depth_v2/ --max_depth_eval 10.0

for KITTI

$ python ./code/eval_with_pngs.py --dataset kitti --split eigen_benchmark --pred_path ./best_kitti_preds/ --gt_path ./datasets/kitti/ --max_depth_eval 80.0 --garg_crop

Evaluate with model (NYU Depth V2)

Result images will be saved in ./args.result_dir/args.exp_name (default: ./results/test)

To evaluate only

$ python ./code/test.py --dataset nyudepthv2 --data_path ./datasets/ --ckpt_dir 
       
         --do_evaluate  --max_depth 10.0 --max_depth_eval 10.0

To save pngs for eval_with_pngs

$ python ./code/test.py --dataset nyudepthv2 --data_path ./datasets/ --ckpt_dir 
       
         --save_eval_pngs  --max_depth 10.0 --max_depth_eval 10.0

To save visualized depth maps

$ python ./code/test.py --dataset nyudepthv2 --data_path ./datasets/ --ckpt_dir 
       
         --save_visualize  --max_depth 10.0 --max_depth_eval 10.0

In case of kitti, modify arguments to --dataset kitti --max_depth 80.0 --max_depth_eval 80.0 and add --kitti_crop [garg_crop or eigen_crop]

Inference

Inference with image directory

$ python ./code/test.py --dataset imagepath --data_path 
     
       --save_visualize

To-Do

Add inference
Add training codes
Add dockerHub link
Add colab

References

[1] From Big to Small: Multi-Scale Local Planar Guidance for Monocular Depth Estimation. [code]

[2] SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers. [code]

Global-Local Path Networks for Monocular Depth Estimation with Vertical CutDepth [Paper]

Related tags

Overview

Global-Local Path Networks for Monocular Depth Estimation with Vertical CutDepth [Paper]

Downloads

Requirements

Inference and Evaluate

Dataset

NYU Depth V2

KITTI

Evaluation

Inference

To-Do

References

Owner

A library to inspect itermediate layers of PyTorch models.

The Most Efficient Temporal Difference Learning Framework for 2048

Improving Compound Activity Classification via Deep Transfer and Representation Learning

Seach Losses of our paper 'Loss Function Discovery for Object Detection via Convergence-Simulation Driven Search', accepted by ICLR 2021.

A Python framework for conversational search

Stacs-ci - A set of modules to enable integration of STACS with commonly used CI / CD systems

CV backbones including GhostNet, TinyNet and TNT, developed by Huawei Noah's Ark Lab.

Real-time LIDAR-based Urban Road and Sidewalk detection for Autonomous Vehicles 🚗

A Multi-modal Perception Tracker (MPT) for speaker tracking using both audio and visual modalities

LVI-SAM: Tightly-coupled Lidar-Visual-Inertial Odometry via Smoothing and Mapping

VISSL is FAIR's library of extensible, modular and scalable components for SOTA Self-Supervised Learning with images.

Rendering color and depth images for ShapeNet models.

Boundary-aware Transformers for Skin Lesion Segmentation

DiAne is a smart fuzzer for IoT devices

Mix3D: Out-of-Context Data Augmentation for 3D Scenes (3DV 2021)

Deep Occlusion-Aware Instance Segmentation with Overlapping BiLayers [CVPR 2021]

Neighbor2Seq: Deep Learning on Massive Graphs by Transforming Neighbors to Sequences

Official PyTorch implementation of PICCOLO: Point-Cloud Centric Omnidirectional Localization (ICCV 2021)

Code for the ICME 2021 paper "Exploring Driving-Aware Salient Object Detection via Knowledge Transfer"