Home

Awesome

FSRE-Depth

This is a Python3 / PyTorch implementation of FSRE-Depth, as described in the following paper:

Fine-grained Semantics-aware Representation Enhancement for Self-supervisedMonocular Depth Estimation overview Hyunyoung Jung, Eunhyeok Park and Sungjoo Yoo

ICCV 2021 (oral)

arXiv pdf

The code was implemented based on Monodepth2.

Setup

This code was implemented under torch==1.3.0 and torchvision==0.4.1, using two NVIDIA TITAN Xp gpus with distrutibted training. Different version may produce different results.

pip install -r requirements.txt

Dataset

KITTI Raw Data and pre-computed segmentation images are required for training.

KITTI/
    ├── 2011_09_26/             
    ├── 2011_09_28/                    
    ├── 2011_09_29/
    ├── 2011_09_30/
    ├── 2011_10_03/
    └── segmentation/   # download and unzip "segmentation.zip" 

Training

For training the full model, run the command as below:

CUDA_VISIBLE_DEVICES=0,1 python -m torch.distributed.launch --nproc_per_node 2 --master_port YOUR_PORT_NUMBER train_ddp.py --data_path YOUR_KITTI_DATA_PATH

Evaluation

The ground truth depth maps should be prepared prior to evaluation.

python export_gt_depth.py --data_path YOUR_KITTI_DATA_PATH --split eigen

MODEL_DIR should be configured as below:

MODEL_DIR
    ├── encoder.pth  # required      
    ├── decoder.pth  # required             
    ├── ...

Run the evaluation command.

python evaluate_depth.py --load_weights_folder MODEL_DIR --data_path YOUR_KITTI_DATA_PATH

Download Models

BackboneInputDownloadAbsRelSqRelRmsRmsLogdelta < 1.25delta < 1.25^2delta < 1.25^3
ResNet-18192 x 640Drive (.zip)0.1050.7084.5460.1820.8860.9640.983

Reference

Please use the following citation when referencing our work:

@InProceedings{Jung_2021_ICCV,
    author    = {Jung, Hyunyoung and Park, Eunhyeok and Yoo, Sungjoo},
    title     = {Fine-Grained Semantics-Aware Representation Enhancement for Self-Supervised Monocular Depth Estimation},
    booktitle = {Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV)},
    month     = {October},
    year      = {2021},
    pages     = {12642-12652}
}