DPDFormer: A Coarse-to-Fine Model for Monocular Depth Estimation

被引:1
|
作者
Liu, Chunpu [1 ,2 ]
Yang, Guanglei [1 ,2 ]
Zuo, Wangmeng [1 ,2 ]
Zang, Tianyi [1 ,2 ]
机构
[1] Harbin Inst Technol, Harbin, Peoples R China
[2] Harbin Inst Technol, Harbin 150001, Peoples R China
关键词
Monocular depth estimation; coarse-to-fine; depth post-discretization; DPDformer;
D O I
10.1145/3638559
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
Monocular depth estimation attracts great attention from computer vision researchers for its convenience in acquiring environment depth information. Recently classification-based MDE methods show its promising performance and begin to act as an essential role in many multi-view applications such as reconstruction and 3D object detection. However, existed classification-based MDE models usually apply fixed depth range discretization strategy across a whole scene. This fixed depth range discretization leads to the imbalance of discretization scale among different depth ranges, resulting in the inexact depth range localization. In this article, to alleviate the imbalanced depth range discretization problem in classification-based monocular depth estimation (MDE) method we follow the coarse-to-fine principle and propose a novel depth range discretization method called depth post-discretization (DPD). Based on a coarse depth anchor roughly indicating the depth range, the DPD generates the depth range discretization adaptively for every position. The depth range discretization with DPD is more fine-grained around the actual depth, which is beneficial for locating the depth range more precisely for each scene position. Besides, to better manage the prediction of the coarse depth anchor and depth probability distribution for calculating the final depth, we design a dual-decoder transformer-based network, i.e., DPDFormer, which is more compatible with our proposed DPD method. We evaluate DPDFormer on popular depth datasets NYU Depth V2 and KITTI. The experimental results prove the superior performance of our proposed method.
引用
收藏
页数:21
相关论文
共 50 条
  • [1] Chfnet: a coarse-to-fine hierarchical refinement model for monocular depth estimation
    Chen, Han
    Wang, Yongxiong
    [J]. MACHINE VISION AND APPLICATIONS, 2024, 35 (04)
  • [2] Coarse-to-fine Planar Regularization for Dense Monocular Depth Estimation
    Liwicki, Stephan
    Zach, Christopher
    Miksik, Ondrej
    Torr, Philip H. S.
    [J]. COMPUTER VISION - ECCV 2016, PT II, 2016, 9906 : 458 - 474
  • [3] Efficient Monocular Coarse-to-Fine Object Pose Estimation
    Feng, Rong
    Zhang, Hong
    [J]. 2016 IEEE INTERNATIONAL CONFERENCE ON MECHATRONICS AND AUTOMATION, 2016, : 1617 - 1622
  • [4] Self-supervised coarse-to-fine monocular depth estimation using a lightweight attention module
    Yuanzhen Li
    Fei Luo
    Chunxia Xiao
    [J]. Computational Visual Media, 2022, 8 : 631 - 647
  • [5] Self-supervised coarse-to-fine monocular depth estimation using a lightweight attention module
    Li, Yuanzhen
    Luo, Fei
    Xiao, Chunxia
    [J]. COMPUTATIONAL VISUAL MEDIA, 2022, 8 (04) : 631 - 647
  • [6] Learning Occlusion-aware Coarse-to-Fine Depth Map for Self-supervised Monocular Depth Estimation
    Zhou, Zhengming
    Dong, Qiulei
    [J]. PROCEEDINGS OF THE 30TH ACM INTERNATIONAL CONFERENCE ON MULTIMEDIA, MM 2022, 2022, : 6386 - 6395
  • [7] Sparse-to-dense coarse-to-fine depth estimation for colonoscopy*
    Liu, Ruyu
    Liu, Zhengzhe
    Lu, Jiaming
    Zhang, Guodao
    Zuo, Zhigui
    Sun, Bo
    Zhang, Jianhua
    Sheng, Weiguo
    Guo, Ran
    Zhang, Lejun
    Hua, Xiaozhen
    [J]. COMPUTERS IN BIOLOGY AND MEDICINE, 2023, 160
  • [8] Face Alignment by Coarse-to-Fine Shape Estimation
    WAN Jun
    LI Jing
    CHANG Jun
    WU Yujia
    XIAO Yafu
    SONG Chengfang
    [J]. Chinese Journal of Electronics, 2018, 27 (06) : 1183 - 1191
  • [9] Coarse-to-fine Animal Pose and Shape Estimation
    Li, Chen
    Lee, Gim Hee
    [J]. ADVANCES IN NEURAL INFORMATION PROCESSING SYSTEMS 34 (NEURIPS 2021), 2021, 34
  • [10] Face Alignment by Coarse-to-Fine Shape Estimation
    Wan Jun
    Li Jing
    Chang Jun
    Wu Yujia
    Xiao Yafu
    Song Chengfang
    [J]. CHINESE JOURNAL OF ELECTRONICS, 2018, 27 (06) : 1183 - 1191