Monocular Semantic Occupancy Grid Mapping With Convolutional Variational Encoder-Decoder Networks

被引：128

作者：

Lu, Chenyang ^{[1
]}

van de Molengraft, Marinus Jacobus Gerardus ^{[2
]}

Dubbelman, Gijs ^{[1
]}

机构：

[1] Eindhoven Univ Technol, Mobile Percept Syst Res Cluster, SPS VCA Grp, Dept Elect Engn, NL-5600 MB Eindhoven, Netherlands

[2] Eindhoven Univ Technol, Dept Mech Engn, Control Syst Technol Grp, NL-5600 MB Eindhoven, Netherlands

来源：

IEEE ROBOTICS AND AUTOMATION LETTERS | 2019年 / 4卷 / 02期

关键词：

Semantic scene understanding; object detection; segmentation and categorization; computer vision for transportation; VISION;

D O I：

10.1109/LRA.2019.2891028

中图分类号：

TP24 [机器人技术];

学科分类号：

080202 ; 1405 ;

摘要：

In this letter, we research and evaluate end-to-end learning of monocular semantic-metric occupancy grid mapping from weak binocular ground truth. The network learns to predict four classes, as well as a camera to bird's eye view mapping. At the core, it utilizes a variational encoder-decoder network that encodes the front-view visual information of the driving scene and subsequently decodes it into a two-dimensional top-view Cartesian coordinate system. The evaluations on Cityscapes show that the end-to-end learning of semantic-metric occupancy grids outperforms the deterministic mapping approach with flat-plane assumption by more than 12% mean intersection-over-union. Furthermore, we show that the variational sampling with a relatively small embedding vector brings robustness against vehicle dynamic perturbations, and generalizability for unseen KITTI data. Our network achieves real-time inference rates of approx. 35 Hz for an input image with a resolution of 256 x 512 pixels and an output map with 64 x 64 occupancy grid cells using a Titan V GPU.

引用

页码：445 / 452

页数：8

共 39 条

[1]

[Anonymous], 2016, P ADV NEUR INF PROC

[2]

[Anonymous], P 3 INT C LEARNING R

[3]

[Anonymous], PROC CVPR IEEE

[4]

[Anonymous], 2017, IEEE I CONF COMP VIS, DOI DOI 10.1109/ICCV.2017.322

[5]

[Anonymous], 2017, COMMUN ACM, DOI DOI 10.1145/3065386

[6]

[Anonymous], 1990, P 6 C UNCERTAINTY AR

[7]

[Anonymous], IEEE T PATTERN ANAL

[8]

[Anonymous], PROC CVPR IEEE

[9]

[Anonymous], ADV NEURAL INFORM PR, DOI DOI 10.1109/TPAMI.2016.2577031

[10]

[Anonymous], 2017, IEEE T PATTERN ANAL, DOI DOI 10.1109/TPAMI.2016.2644615

← 1 2 3 4 →