DPDFormer: A Coarse-to-Fine Model for Monocular Depth Estimation

被引：1

作者：

Liu, Chunpu ^{[1
,2
]}

Yang, Guanglei ^{[1
,2
]}

Zuo, Wangmeng ^{[1
,2
]}

Zang, Tianyi ^{[1
,2
]}

机构：

[1] Harbin Inst Technol, Harbin, Peoples R China

[2] Harbin Inst Technol, Harbin 150001, Peoples R China

来源：

ACM TRANSACTIONS ON MULTIMEDIA COMPUTING COMMUNICATIONS AND APPLICATIONS | 2024年 / 20卷 / 05期

关键词：

Monocular depth estimation; coarse-to-fine; depth post-discretization; DPDformer;

D O I：

10.1145/3638559

中图分类号：

TP [自动化技术、计算机技术];

学科分类号：

0812 ;

摘要：

Monocular depth estimation attracts great attention from computer vision researchers for its convenience in acquiring environment depth information. Recently classification-based MDE methods show its promising performance and begin to act as an essential role in many multi-view applications such as reconstruction and 3D object detection. However, existed classification-based MDE models usually apply fixed depth range discretization strategy across a whole scene. This fixed depth range discretization leads to the imbalance of discretization scale among different depth ranges, resulting in the inexact depth range localization. In this article, to alleviate the imbalanced depth range discretization problem in classification-based monocular depth estimation (MDE) method we follow the coarse-to-fine principle and propose a novel depth range discretization method called depth post-discretization (DPD). Based on a coarse depth anchor roughly indicating the depth range, the DPD generates the depth range discretization adaptively for every position. The depth range discretization with DPD is more fine-grained around the actual depth, which is beneficial for locating the depth range more precisely for each scene position. Besides, to better manage the prediction of the coarse depth anchor and depth probability distribution for calculating the final depth, we design a dual-decoder transformer-based network, i.e., DPDFormer, which is more compatible with our proposed DPD method. We evaluate DPDFormer on popular depth datasets NYU Depth V2 and KITTI. The experimental results prove the superior performance of our proposed method.

引用

页数：21

共 50 条

[1] Chfnet: a coarse-to-fine hierarchical refinement model for monocular depth estimation
Chen, Han
Wang, Yongxiong
[J]. MACHINE VISION AND APPLICATIONS, 2024, 35 (04)
[2] Coarse-to-fine Planar Regularization for Dense Monocular Depth Estimation
Liwicki, Stephan
Zach, Christopher
Miksik, Ondrej
Torr, Philip H. S.
[J]. COMPUTER VISION - ECCV 2016, PT II, 2016, 9906 : 458 - 474
[3] Efficient Monocular Coarse-to-Fine Object Pose Estimation
Feng, Rong
Zhang, Hong
[J]. 2016 IEEE INTERNATIONAL CONFERENCE ON MECHATRONICS AND AUTOMATION, 2016, : 1617 - 1622
[4] Self-supervised coarse-to-fine monocular depth estimation using a lightweight attention module
Yuanzhen Li
Fei Luo
Chunxia Xiao
[J]. Computational Visual Media, 2022, 8 : 631 - 647
[5] Self-supervised coarse-to-fine monocular depth estimation using a lightweight attention module
Li, Yuanzhen
Luo, Fei
Xiao, Chunxia
[J]. COMPUTATIONAL VISUAL MEDIA, 2022, 8 (04) : 631 - 647
[6] Learning Occlusion-aware Coarse-to-Fine Depth Map for Self-supervised Monocular Depth Estimation
Zhou, Zhengming
Dong, Qiulei
[J]. PROCEEDINGS OF THE 30TH ACM INTERNATIONAL CONFERENCE ON MULTIMEDIA, MM 2022, 2022, : 6386 - 6395
[7] Sparse-to-dense coarse-to-fine depth estimation for colonoscopy*
Liu, Ruyu
Liu, Zhengzhe
Lu, Jiaming
Zhang, Guodao
Zuo, Zhigui
Sun, Bo
Zhang, Jianhua
Sheng, Weiguo
Guo, Ran
Zhang, Lejun
Hua, Xiaozhen
[J]. COMPUTERS IN BIOLOGY AND MEDICINE, 2023, 160
[8] Face Alignment by Coarse-to-Fine Shape Estimation
WAN Jun
LI Jing
CHANG Jun
WU Yujia
XIAO Yafu
SONG Chengfang
[J]. Chinese Journal of Electronics, 2018, 27 (06) : 1183 - 1191
[9] Coarse-to-fine Animal Pose and Shape Estimation
Li, Chen
Lee, Gim Hee
[J]. ADVANCES IN NEURAL INFORMATION PROCESSING SYSTEMS 34 (NEURIPS 2021), 2021, 34
[10] Face Alignment by Coarse-to-Fine Shape Estimation
Wan Jun
Li Jing
Chang Jun
Wu Yujia
Xiao Yafu
Song Chengfang
[J]. CHINESE JOURNAL OF ELECTRONICS, 2018, 27 (06) : 1183 - 1191

← 1 2 3 4 5 →