Next Article in Journal
Minimum Noise Fraction Analysis of TGO/NOMAD LNO Channel High-Resolution Nadir Spectra of Mars
Previous Article in Journal
Combining Satellite Imagery and a Deep Learning Algorithm to Retrieve the Water Levels of Small Reservoirs
Previous Article in Special Issue
RNGC-VIWO: Robust Neural Gyroscope Calibration Aided Visual-Inertial-Wheel Odometry for Autonomous Vehicle
 
 
Font Type:
Arial Georgia Verdana
Font Size:
Aa Aa Aa
Line Spacing:
Column Width:
Background:
Article

Switchable-Encoder-Based Self-Supervised Learning Framework for Monocular Depth and Pose Estimation

1
Department of Multimedia Engineering, Dongguk University-Seoul, 30, Pildongro-1-gil, Jung-gu, Seoul 04620, Republic of Korea
2
Autonomous Driving Research Department, KoROAD (Korea Road Traffic Authority) 2, Hyeoksin-ro, Wonu-si, Gangwon-do 26466, Republic of Korea
3
Division of AI Software Convergence, Dongguk University-Seoul, 30, Pildongro-1-gil, Jung-gu, Seoul 04620, Republic of Korea
*
Author to whom correspondence should be addressed.
Remote Sens. 2023, 15(24), 5739; https://doi.org/10.3390/rs15245739
Submission received: 4 September 2023 / Revised: 29 November 2023 / Accepted: 13 December 2023 / Published: 15 December 2023
(This article belongs to the Special Issue Signal Processing and Machine Learning for Autonomous Vehicles)

Abstract

Monocular depth prediction research is essential for expanding meaning from 2D to 3D. Recent studies have focused on the application of a newly proposed encoder; however, the development within the self-supervised learning framework remains unexplored, an aspect critical for advancing foundational models of 3D semantic interpretation. Addressing the dynamic nature of encoder-based research, especially in performance evaluations for feature extraction and pre-trained models, this research proposes the switchable encoder learning framework (SELF). SELF enhances versatility by enabling the seamless integration of diverse encoders in a self-supervised learning context for depth prediction. This integration is realized through the direct transfer of feature information from the encoder and by standardizing the input structure of the decoder to accommodate various encoder architectures. Furthermore, the framework is extended and incorporated into an adaptable decoder for depth prediction and camera pose learning, employing standard loss functions. Comparative experiments with previous frameworks using the same encoder reveal that SELF achieves a 7% reduction in parameters while enhancing performance. Remarkably, substituting newly proposed algorithms in place of an encoder improves the outcomes as well as significantly decreases the number of parameters by 23%. The experimental findings highlight the ability of SELF to broaden depth factors, such as depth consistency. This framework facilitates the objective selection of algorithms as a backbone for extended research in monocular depth prediction.
Keywords: structure from motion; self-supervised learning; monocular depth estimation structure from motion; self-supervised learning; monocular depth estimation

Share and Cite

MDPI and ACS Style

Kim, J.; Gao, R.; Park, J.; Yoon, J.; Cho, K. Switchable-Encoder-Based Self-Supervised Learning Framework for Monocular Depth and Pose Estimation. Remote Sens. 2023, 15, 5739. https://doi.org/10.3390/rs15245739

AMA Style

Kim J, Gao R, Park J, Yoon J, Cho K. Switchable-Encoder-Based Self-Supervised Learning Framework for Monocular Depth and Pose Estimation. Remote Sensing. 2023; 15(24):5739. https://doi.org/10.3390/rs15245739

Chicago/Turabian Style

Kim, Junoh, Rui Gao, Jisun Park, Jinsoo Yoon, and Kyungeun Cho. 2023. "Switchable-Encoder-Based Self-Supervised Learning Framework for Monocular Depth and Pose Estimation" Remote Sensing 15, no. 24: 5739. https://doi.org/10.3390/rs15245739

APA Style

Kim, J., Gao, R., Park, J., Yoon, J., & Cho, K. (2023). Switchable-Encoder-Based Self-Supervised Learning Framework for Monocular Depth and Pose Estimation. Remote Sensing, 15(24), 5739. https://doi.org/10.3390/rs15245739

Note that from the first issue of 2016, this journal uses article numbers instead of page numbers. See further details here.

Article Metrics

Back to TopTop