IROS 2021poster2 citations

Improving Monocular Depth Estimation by Semantic Pre-training

Peter Rottmann, Thorbjörn Posewsky, Andres Milioto, Cyrill Stachniss, Jens Behley

Abstract

Knowing the distance to nearby objects is crucial for autonomous cars to navigate safely in everyday traffic. In this paper, we investigate monocular depth estimation, which advanced substantially within the last years and is providing increasingly more accurate results while only requiring a single camera image as input. In line with recent work, we use an encoder-decoder structure with so-called packing layers to estimate depth values in a self-supervised fashion. We propose integrating a joint pre-training of semantic segmentation plus depth estimation on a dataset providing semantic labels. By using a separate semantic decoder that is only needed for pre-training, we can keep the network comparatively small. Our extensive experimental evaluation shows that the addition of such pre-training improves the depth estimation performance substantially. Finally, we show that we achieve competitive performance on the KITTI dataset despite using a much smaller and more efficient network.

BibTeX
@inproceedings{iros2021_improvingmonocul,
  title = {Improving Monocular Depth Estimation by Semantic Pre-training},
  author = {Peter Rottmann and Thorbjörn Posewsky and Andres Milioto and Cyrill Stachniss and Jens Behley},
  booktitle = {IROS 2021},
  year = {2021}
}
Improving Monocular Depth Estimation by Semantic Pre-training · IROS 2021