Effective fusion module with dilation convolution for monocular panoramic depth estimate

Cheng Han; Yongqing Cai; Xinpeng Pan; Ziyun Wang

doi:10.1049/ipr2.13007

IET Image Processing (Mar 2024)

Effective fusion module with dilation convolution for monocular panoramic depth estimate

Cheng Han,
Yongqing Cai,
Xinpeng Pan,
Ziyun Wang

Affiliations

Cheng Han: School of Computer Science and Technology Changchun University of Science and Technology Changchun China
Yongqing Cai: School of Computer Science and Technology Changchun University of Science and Technology Changchun China
Xinpeng Pan: School of Computer Science and Technology Changchun University of Science and Technology Changchun China
Ziyun Wang: School of Computer Science and Technology Changchun University of Science and Technology Changchun China

DOI: https://doi.org/10.1049/ipr2.13007
Journal volume & issue: Vol. 18, no. 4
pp. 1073 – 1082

Abstract

Read online

Abstract Depth estimation from monocular panoramic image is a crucial step in 3D reconstruction, which is a close relationship with virtual reality and metaverse technologies. In recent years, some methods, such as HRDFuse, BiFuse++, and UniFuse, have employed a two‐branch neural network leveraging two common projections: equirectangular and cubemap projections (CMPs). The equirectangular projection (ERP) provides a complete field of view but introduces distortion, while the CMP avoids distortion but introduces discontinuity at the boundary of the cube. In order to address the issue of distortion and discontinuity, the authors propose an efficient depth estimation fusion module to balance the feature mapping of the two projections. Moreover, for the ERP, the authors propose a novel inflated network architecture to extend the receptive field and effectively harness visual information. Extensive experiments show that the authors’ method predicts more clear boundaries and accurate depth results while outperforming mainstream panoramic depth estimation algorithms.

Published in IET Image Processing

ISSN: 1751-9659 (Print); 1751-9667 (Online)
Publisher: Wiley
Country of publisher: United Kingdom
LCC subjects: Technology: Photography; Science: Mathematics: Instruments and machines: Electronic computers. Computer science: Computer software
Website: https://ietresearch.onlinelibrary.wiley.com/journal/17519667

About the journal

Abstract

Keywords