基本事实
航程(航空)
计算机科学
人工智能
卷积神经网络
深度学习
单眼
RGB颜色模型
计算机视觉
比例(比率)
深度图
生成对抗网络
人工神经网络
图像(数学)
地理
地图学
工程类
航空航天工程
作者
Md. Alimoor Reza,Jana Košecká,Philip David
标识
DOI:10.1109/iros.2018.8593971
摘要
This paper introduces the problem of long-range monocular depth estimation for outdoor urban environments. Range sensors and traditional depth estimation algorithms (both stereo and single view) predict depth for distances of less than 100 meters in outdoor settings and 10 meters in indoor settings. The shortcomings of outdoor single view methods that use learning approaches are, to some extent, due to the lack of long-range ground truth training data, which in turn is due to limitations of range sensors. To circumvent this, we first propose a novel strategy for generating synthetic long-range ground truth depth data. We utilize Google Earth images to reconstruct large-scale 3D models of different cities with proper scale. The acquired repository of 3D models and associated RGB views along with their long-range depth renderings are used as training data for depth prediction. We then train two deep neural network models for long-range depth estimation: i) a Convolutional Neural Network (CNN) and ii) a Generative Adversarial Network (GAN). We found in our experiments that the GAN model predicts depth more accurately. We plan to open-source the database and the baseline models for public use.
科研通智能强力驱动
Strongly Powered by AbleSci AI