GitHub - msdkhairi/DPT: Dense Prediction Transformers

Evaluation V2

python run_monodepth.py --input_path="/home/data/kitti" --data_filenames_path="eigen_benchmark/test_list.txt" --output_path="output_monodepth" --model_type=dpt_hybrid_kitti --kitti_crop --absolute_depth --no-optimize

python ./eval_with_pngs.py --pred_path ./output_monodepth/ --gt_path /home/data/kitti --dataset kitti --min_depth_eval 1e-3 --max_depth_eval 80 --garg_crop --do_kb_crop --data_filenames_path="eigen_benchmark/test_list.txt"

Vision Transformers for Dense Prediction

This repository contains code and models for our paper:

Vision Transformers for Dense Prediction
René Ranftl, Alexey Bochkovskiy, Vladlen Koltun

Changelog

[March 2021] Initial release of inference code and models

Setup

Download the model weights and place them in the weights folder:

Monodepth:

Segmentation:

Set up dependencies:
```
pip install -r requirements.txt
```
The code was tested with Python 3.7, PyTorch 1.8.0, OpenCV 4.5.1, and timm 0.4.5

Usage

Place one or more input images in the folder input.
Run a monocular depth estimation model:
```
python run_monodepth.py
```
Or run a semantic segmentation model:
```
python run_segmentation.py
```
The results are written to the folder output_monodepth and output_semseg, respectively.

Use the flag -t to switch between different models. Possible options are dpt_hybrid (default) and dpt_large.

Additional models:

Monodepth finetuned on KITTI: dpt_hybrid_kitti-cb926ef4.pt Mirror
Monodepth finetuned on NYUv2: dpt_hybrid_nyu-2ce69ec7.pt Mirror

Run with

python run_monodepth -t [dpt_hybrid_kitti|dpt_hybrid_nyu]

Evaluation

Hints on how to evaluate monodepth models can be found here: https://github.com/intel-isl/DPT/blob/main/EVALUATION.md

Citation

Please cite our papers if you use this code or any of the models.

@article{Ranftl2021,
	author    = {Ren\'{e} Ranftl and Alexey Bochkovskiy and Vladlen Koltun},
	title     = {Vision Transformers for Dense Prediction},
	journal   = {ArXiv preprint},
	year      = {2021},
}

@article{Ranftl2020,
	author    = {Ren\'{e} Ranftl and Katrin Lasinger and David Hafner and Konrad Schindler and Vladlen Koltun},
	title     = {Towards Robust Monocular Depth Estimation: Mixing Datasets for Zero-shot Cross-dataset Transfer},
	journal   = {IEEE Transactions on Pattern Analysis and Machine Intelligence (TPAMI)},
	year      = {2020},
}

Acknowledgements

Our work builds on and uses code from timm and PyTorch-Encoding. We'd like to thank the authors for making these libraries available.

License

MIT License

Name		Name	Last commit message	Last commit date
Latest commit History 128 Commits
dpt		dpt
eigen_benchmark		eigen_benchmark
input		input
output_semseg		output_semseg
util		util
weights		weights
.gitignore		.gitignore
EVALUATION.md		EVALUATION.md
LICENSE		LICENSE
README.md		README.md
environment.yml		environment.yml
eval_with_pngs.py		eval_with_pngs.py
playground.ipynb		playground.ipynb
requirements.txt		requirements.txt
run_monodepth.py		run_monodepth.py
run_segmentation.py		run_segmentation.py
setup.py		setup.py
validate_kitti.py		validate_kitti.py

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

Evaluation V2

Vision Transformers for Dense Prediction

Changelog

Setup

Usage

Evaluation

Citation

Acknowledgements

License

About

Uh oh!

Releases

Packages

Uh oh!

Contributors

Uh oh!

Languages

Folders and files

Latest commit

History

Repository files navigation

Evaluation V2

Vision Transformers for Dense Prediction

Changelog

Setup

Usage

Evaluation

Citation

Acknowledgements

License

About

Resources

License

Uh oh!

Stars

Watchers

Forks

Releases

Packages 0

Uh oh!

Contributors

Uh oh!

Languages

Packages