Global-Local Path Networks for Monocular Depth Estimation with Vertical CutDepth [Paper]

Last update: Jan 01, 2023

Related tags

Deep Learning GLPDepth

Overview

Global-Local Path Networks for Monocular Depth Estimation with Vertical CutDepth [Paper]

Downloads

[Downloads] Trained ckpt files for NYU Depth V2 and KITTI
[Downloads] Predicted depth maps png files for NYU Depth V2 and KITTI Eigen split test set

Requirements

Tested on

python==3.7.7
torch==1.6.0
h5py==3.6.0
scipy==1.7.3
opencv-python==4.5.5
mmcv==1.4.3
timm=0.5.4
albumentations=1.1.0
tensorboardX==2.4.1

You can install above package with

$ pip install -r requirements.txt

Inference and Evaluate

Dataset

NYU Depth V2

$ cd ./datasets
$ wget http://horatio.cs.nyu.edu/mit/silberman/nyu_depth_v2/nyu_depth_v2_labeled.mat
$ python ../code/utils/extract_official_train_test_set_from_mat.py nyu_depth_v2_labeled.mat splits.mat ./nyu_depth_v2/official_splits/

KITTI

Download annotated depth maps data set (14GB) from [link] into ./datasets/kitti/data_depth_annotated

$ cd ./datasets/kitti/data_depth_annotated/
$ unzip data_depth_annotated.zip

With above two instrtuctions, you can perform eval_with_pngs.py/test.py for NYU Depth V2 and eval_with_pngs for KITTI.

To fully perform experiments, please follow [BTS] repository to obtain full dataset for NYU Depth V2 and KITTI datasets.

Your dataset directory should be

root
- nyu_depth_v2
  - bathroom_0001
  - bathroom_0002
  - ...
  - official_splits
- kitti
  - data_depth_annotated
  - raw_data
  - val_selection_cropped

Evaluation

Evaluate with png images

for NYU Depth V2

$ python ./code/eval_with_pngs.py --dataset nyudepthv2 --pred_path ./best_nyu_preds/ --gt_path ./datasets/nyu_depth_v2/ --max_depth_eval 10.0

for KITTI

$ python ./code/eval_with_pngs.py --dataset kitti --split eigen_benchmark --pred_path ./best_kitti_preds/ --gt_path ./datasets/kitti/ --max_depth_eval 80.0 --garg_crop

Evaluate with model (NYU Depth V2)

Result images will be saved in ./args.result_dir/args.exp_name (default: ./results/test)

To evaluate only

$ python ./code/test.py --dataset nyudepthv2 --data_path ./datasets/ --ckpt_dir 
       
         --do_evaluate  --max_depth 10.0 --max_depth_eval 10.0

To save pngs for eval_with_pngs

$ python ./code/test.py --dataset nyudepthv2 --data_path ./datasets/ --ckpt_dir 
       
         --save_eval_pngs  --max_depth 10.0 --max_depth_eval 10.0

To save visualized depth maps

$ python ./code/test.py --dataset nyudepthv2 --data_path ./datasets/ --ckpt_dir 
       
         --save_visualize  --max_depth 10.0 --max_depth_eval 10.0

In case of kitti, modify arguments to --dataset kitti --max_depth 80.0 --max_depth_eval 80.0 and add --kitti_crop [garg_crop or eigen_crop]

Inference

Inference with image directory

$ python ./code/test.py --dataset imagepath --data_path 
     
       --save_visualize

To-Do

Add inference
Add training codes
Add dockerHub link
Add colab

References

[1] From Big to Small: Multi-Scale Local Planar Guidance for Monocular Depth Estimation. [code]

[2] SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers. [code]

Global-Local Path Networks for Monocular Depth Estimation with Vertical CutDepth [Paper]

Related tags

Overview

Global-Local Path Networks for Monocular Depth Estimation with Vertical CutDepth [Paper]

Downloads

Requirements

Inference and Evaluate

Dataset

NYU Depth V2

KITTI

Evaluation

Inference

To-Do

References

Owner

Code for "Diversity can be Transferred: Output Diversification for White- and Black-box Attacks"

Recommendation algorithms for large graphs

DVG-Face: Dual Variational Generation for Heterogeneous Face Recognition, TPAMI 2021

EqGAN - Improving GAN Equilibrium by Raising Spatial Awareness

Repository to run object detection on a model trained on an autonomous driving dataset.

CNNs for Sentence Classification in PyTorch

A certifiable defense against adversarial examples by training neural networks to be provably robust

Bling's Object detection tool

Real-time pose estimation accelerated with NVIDIA TensorRT

Datasets, Transforms and Models specific to Computer Vision

Regularizing Nighttime Weirdness: Efficient Self-supervised Monocular Depth Estimation in the Dark (ICCV 2021)

UnsupervisedR&R: Unsupervised Pointcloud Registration via Differentiable Rendering

Code for "Continuous-Time Meta-Learning with Forward Mode Differentiation" (ICLR 2022)

PyTorch implementation of Masked Autoencoders Are Scalable Vision Learners for self-supervised ViT.

ECCV18 Workshops - Enhanced SRGAN. Champion PIRM Challenge on Perceptual Super-Resolution. The training codes are in BasicSR.

[ WSDM '22 ] On Sampling Collaborative Filtering Datasets

Running AlphaFold2 (from ColabFold) in Azure Machine Learning

Microsoft Cognitive Toolkit (CNTK), an open source deep-learning toolkit

Robust Video Matting in PyTorch, TensorFlow, TensorFlow.js, ONNX, CoreML!

C3d-pytorch - Pytorch porting of C3D network, with Sports1M weights