Single Image Depth Prediction with Wavelet Decomposition

Michael Ramamonjisoa,Michael Firman,Jamie Watson,Vincent Lepetit,Daniyar Turmukhambetov

2021 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION, CVPR 2021（2021）

引用 47|浏览80

暂无评分

摘要

We present a novel method for predicting accurate depths from monocular images with high efficiency. This optimal efficiency is achieved by exploiting wavelet decomposition, which is integrated in a fully differentiable encoder-decoder architecture. We demonstrate that we can reconstruct high-fidelity depth maps by predicting sparse wavelet coefficients. In contrast with previous works, we show that wavelet coefficients can be learned without direct supervision on coefficients. Instead we supervise only the final depth image that is reconstructed through the inverse wavelet transform. We additionally show that wavelet coefficients can be learned in fully self-supervised scenarios, without access to ground-truth depth. Finally, we apply our method to different state-of-the-art monocular depth estimation models, in each case giving similar or better results compared to the original model, while requiring less than half the multiplyadds in the decoder network.

查看译文

关键词

single image depth prediction,wavelet decomposition,accurate depths,monocular images,optimal efficiency,fully differentiable encoder-decoder architecture,high-fidelity depth maps,sparse wavelet coefficients,direct supervision,final depth image,self-supervised scenarios,ground-truth depth,different state-of-the-art monocular depth estimation models

AI 理解论文

溯源树

样例

生成溯源树，研究论文发展脉络

Chat Paper

正在生成论文摘要