跳到主要导航 跳到搜索 跳到主要内容

MLDA-Net: Multi-Level Dual Attention-Based Network for Self-Supervised Monocular Depth Estimation

  • Xibin Song
  • , Wei Li
  • , Dingfu Zhou
  • , Yuchao Dai
  • , Jin Fang
  • , Hongdong Li
  • , Liangjun Zhang
  • Baidu Inc
  • Shandong University
  • Australian National University

科研成果: 期刊稿件文章同行评审

52 引用 (Scopus)

摘要

The success of supervised learning-based single image depth estimation methods critically depends on the availability of large-scale dense per-pixel depth annotations, which requires both laborious and expensive annotation process. Therefore, the self-supervised methods are much desirable, which attract significant attention recently. However, depth maps predicted by existing self-supervised methods tend to be blurry with many depth details lost. To overcome these limitations, we propose a novel framework, named MLDA-Net, to obtain per-pixel depth maps with shaper boundaries and richer depth details. Our first innovation is a multi-level feature extraction (MLFE) strategy which can learn rich hierarchical representation. Then, a dual-attention strategy, combining global attention and structure attention, is proposed to intensify the obtained features both globally and locally, resulting in improved depth maps with sharper boundaries. Finally, a reweighted loss strategy based on multi-level outputs is proposed to conduct effective supervision for self-supervised depth estimation. Experimental results demonstrate that our MLDA-Net framework achieves state-of-the-art depth prediction results on the KITTI benchmark for self-supervised monocular depth estimation with different input modes and training modes. Extensive experiments on other benchmark datasets further confirm the superiority of our proposed approach.

源语言英语
期刊论文编号9416235
页(从-至)4691-4705
页数15
期刊IEEE Transactions on Image Processing
30
DOI
出版状态已出版 - 2021

学术指纹

探究 'MLDA-Net: Multi-Level Dual Attention-Based Network for Self-Supervised Monocular Depth Estimation' 的科研主题。它们共同构成独一无二的学术指纹。

引用此