Rectifying Pseudo Label Learning via Uncertainty Estimation for Domain Adaptive Semantic Segmentation

被引:354
作者
Zheng, Zhedong [1 ]
Yang, Yi [1 ]
机构
[1] Univ Technol Sydney, Australian Artificial Intelligence Inst AAII, Ultimo, NSW, Australia
关键词
Unsupervised domain adaptation; Domain adaptive semantic segmentation; Image segmentation; Uncertainty estimation;
D O I
10.1007/s11263-020-01395-y
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
This paper focuses on the unsupervised domain adaptation of transferring the knowledge from the source domain to the target domain in the context of semantic segmentation. Existing approaches usually regard the pseudo label as the ground truth to fully exploit the unlabeled target-domain data. Yet the pseudo labels of the target-domain data are usually predicted by the model trained on the source domain. Thus, the generated labels inevitably contain the incorrect prediction due to the discrepancy between the training domain and the test domain, which could be transferred to the final adapted model and largely compromises the training process. To overcome the problem, this paper proposes to explicitly estimate the prediction uncertainty during training to rectify the pseudo label learning for unsupervised semantic segmentation adaptation. Given the input image, the model outputs the semantic segmentation prediction as well as the uncertainty of the prediction. Specifically, we model the uncertainty via the prediction variance and involve the uncertainty into the optimization objective. To verify the effectiveness of the proposed method, we evaluate the proposed method on two prevalent synthetic-to-real semantic segmentation benchmarks, i.e., GTA5 -> Cityscapes and SYNTHIA -> Cityscapes, as well as one cross-city benchmark, i.e., Cityscapes -> Oxford RobotCar. We demonstrate through extensive experiments that the proposed approach (1) dynamically sets different confidence thresholds according to the prediction variance, (2) rectifies the learning from noisy pseudo labels, and (3) achieves significant improvements over the conventional pseudo label learning and yields competitive performance on all three benchmarks.
引用
收藏
页码:1106 / 1120
页数:15
相关论文
共 52 条
  • [31] Maximum Classifier Discrepancy for Unsupervised Domain Adaptation
    Saito, Kuniaki
    Watanabe, Kohei
    Ushiku, Yoshitaka
    Harada, Tatsuya
    [J]. 2018 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2018, : 3723 - 3732
  • [32] Srivastava N, 2014, J MACH LEARN RES, V15, P1929
  • [33] Domain Adaptation for Structured Output via Discriminative Patch Representations
    Tsai, Yi-Hsuan
    Sohn, Kihyuk
    Schulter, Samuel
    Chandraker, Manmohan
    [J]. 2019 IEEE/CVF INTERNATIONAL CONFERENCE ON COMPUTER VISION (ICCV 2019), 2019, : 1456 - 1465
  • [34] Learning to Adapt Structured Output Space for Semantic Segmentation
    Tsai, Yi-Hsuan
    Hung, Wei-Chih
    Schulter, Samuel
    Sohn, Kihyuk
    Yang, Ming-Hsuan
    Chandraker, Manmohan
    [J]. 2018 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2018, : 7472 - 7481
  • [35] ADVENT: Adversarial Entropy Minimization for Domain Adaptation in Semantic Segmentation
    Tuan-Hung Vu
    Jain, Himalaya
    Bucher, Maxime
    Cord, Matthieu
    Perez, Patrick
    [J]. 2019 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2019), 2019, : 2512 - 2521
  • [36] Simultaneous Deep Transfer Across Domains and Tasks
    Tzeng, Eric
    Hoffman, Judy
    Darrell, Trevor
    Saenko, Kate
    [J]. 2015 IEEE INTERNATIONAL CONFERENCE ON COMPUTER VISION (ICCV), 2015, : 4068 - 4076
  • [37] Transferable Joint Attribute-Identity Deep Learning for Unsupervised Person Re-Identification
    Wang, Jingya
    Zhu, Xiatian
    Gong, Shaogang
    Li, Wei
    [J]. 2018 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2018, : 2275 - 2284
  • [38] Class-Specific Reconstruction Transfer Learning for Visual Recognition Across Domains
    Wang, Shanshan
    Zhang, Lei
    Zuo, Wangmeng
    Zhang, Bob
    [J]. IEEE TRANSACTIONS ON IMAGE PROCESSING, 2020, 29 : 2424 - 2438
  • [39] Revisiting Dilated Convolution: A Simple Approach for Weakly- and Semi-Supervised Semantic Segmentation
    Wei, Yunchao
    Xiao, Huaxin
    Shi, Honghui
    Jie, Zequn
    Feng, Jiashi
    Huang, Thomas S.
    [J]. 2018 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2018, : 7268 - 7277
  • [40] DCAN: Dual Channel-Wise Alignment Networks for Unsupervised Scene Adaptation
    Wu, Zuxuan
    Han, Xintong
    Lin, Yen-Liang
    Uzunbas, Mustafa Gokhan
    Goldstein, Tom
    Lim, Ser Nam
    Davis, Larry S.
    [J]. COMPUTER VISION - ECCV 2018, PT V, 2018, 11209 : 535 - 552