Bird's-Eye-View Panoptic Segmentation Using Monocular Frontal View Images

被引:35
|
作者
Gosala, Nikhil [1 ]
Valada, Abhinav [1 ]
机构
[1] Univ Freiburg, Dept Comp Sci, Freiburg, Germany
关键词
Semantic scene understanding; object detection; segmentation and categorization; deep learning for visual perception;
D O I
10.1109/LRA.2022.3142418
中图分类号
TP24 [机器人技术];
学科分类号
080202 ; 1405 ;
摘要
Bird's-Eye-View (BEV) maps have emerged as one of the most powerful representations for scene understanding due to their ability to provide rich spatial context while being easy to interpret and process. Such maps have found use in many real-world tasks that extensively rely on accurate scene segmentation as well as object instance identification in the BEV space for their operation. However, existing segmentation algorithms only predict the semantics in the BEV space, which limits their use in applications where the notion of object instances is also critical. In this work, we present the first BEV panoptic segmentation approach for directly predicting dense panoptic segmentation maps in the BEV, given a single monocular image in the frontal view (FV). Our architecture follows the top-down paradigm and incorporates a novel dense transformer module consisting of two distinct transformers that learn to independently map vertical and flat regions in the input image from the FVto the BEV. Additionally, we derive a mathematical formulation for the sensitivity of the FV-BEV transformation which allows us to intelligently weight pixels in the BEV space to account for the varying descriptiveness across the FV image. Extensive evaluations on the KITTI-360 and nuScenes datasets demonstrate that our approach exceeds the state-of-the-art in the PQ metric by 3.61 pp and 4.93 pp respectively.
引用
收藏
页码:1968 / 1975
页数:8
相关论文
共 50 条
  • [1] SkyEye: Self-Supervised Bird's-Eye-View Semantic Mapping Using Monocular Frontal View Images
    Gosala, Nikhil
    Petek, Kuersat
    Drews-, Paulo L. J., Jr.
    Burgard, Wolfram
    Valada, Abhinav
    2023 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2023, : 14901 - 14910
  • [2] Camera-view supervision for bird's-eye-view semantic segmentation
    Yang, Bowen
    Yu, Linlin
    Chen, Feng
    FRONTIERS IN BIG DATA, 2024, 7
  • [3] 3D Bird's-Eye-View Instance Segmentation
    Elich, Cathrin
    Engelmann, Francis
    Kontogianni, Theodora
    Leibe, Bastian
    PATTERN RECOGNITION, DAGM GCPR 2019, 2019, 11824 : 48 - 61
  • [4] BEVSeg: Geometry and Data-Driven Based Multi-view Segmentation in Bird's-Eye-View
    Chen, Qiuxiao
    Tai, Hung-Shuo
    Li, Pengfei
    Wang, Ke
    Qi, Xiaojun
    COMPUTER VISION SYSTEMS, ICVS 2023, 2023, 14253 : 432 - 443
  • [5] Structured Bird's-Eye-View Traffic Scene Understanding from Onboard Images
    Can, Yigit Baran
    Liniger, Alexander
    Paudel, Danda Pani
    Van Gool, Luc
    2021 IEEE/CVF INTERNATIONAL CONFERENCE ON COMPUTER VISION (ICCV 2021), 2021, : 15641 - 15650
  • [6] X-Align++: cross-modal cross-view alignment for Bird’s-eye-view segmentation
    Shubhankar Borse
    Marvin Klingner
    Varun Ravi
    Hong Cai
    Abdulaziz Almuzairee
    Senthil Yogamani
    Fatih Porikli
    Machine Vision and Applications, 2023, 34
  • [7] X-Align: Cross-Modal Cross-View Alignment for Bird's-Eye-View Segmentation
    Borse, Shubhankar
    Klingner, Marvin
    Kumar, Varun Ravi
    Cai, Hong
    Almuzairee, Abdulaziz
    Yogamani, Senthil
    Porikli, Fatih
    2023 IEEE/CVF WINTER CONFERENCE ON APPLICATIONS OF COMPUTER VISION (WACV), 2023, : 3286 - 3296
  • [8] LaRa: Latents and Rays for Multi-Camera Bird's-Eye-View Semantic Segmentation
    Bartoccioni, Florent
    Zablocki, Eloi
    Bursuc, Andrei
    Perez, Patrick
    Cord, Matthieu
    Alahari, Karteek
    CONFERENCE ON ROBOT LEARNING, VOL 205, 2022, 205 : 1663 - 1672
  • [9] DVT: Decoupled Dual-Branch View Transformation for Monocular Bird's Eye View Semantic Segmentation
    Du, Jiayuan
    Pan, Xianghui
    Shen, Mengjiao
    Su, Shuai
    Yang, Jingwei
    Liu, Chengju
    Chen, Qijun
    2024 IEEE/RSJ INTERNATIONAL CONFERENCE ON INTELLIGENT ROBOTS AND SYSTEMS (IROS 2024), 2024, : 9769 - 9776
  • [10] Predicting Bird's-Eye-View Semantic Representations Using Correlated Context Learning
    Chen, Yongquan
    Fan, Weiming
    Zheng, Wenli
    Huang, Rui
    Yu, Jiahui
    IEEE ROBOTICS AND AUTOMATION LETTERS, 2024, 9 (05): : 4718 - 4725