Guided CNN for generalized zero-shot and open-set recognition using visual and semantic prototypes

被引:32
作者
Geng, Chuanxing [1 ]
Tao, Lue [1 ]
Chen, Songcan [1 ]
机构
[1] Nanjing Univ Aeronaut & Astronaut, Coll Comp Sci & Technol, MIIT Key Lab Pattern Anal & Machine Intelligence, Nanjing 211106, Peoples R China
关键词
Convolutional prototype learning; Generalized zero-shot Learning; Open set recognition;
D O I
10.1016/j.patcog.2020.107263
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
In the process of exploring the world, the curiosity constantly drives humans to cognize new things. Supposing you are a zoologist, for a presented animal image, you can recognize it immediately if you know its class. Otherwise, you would more likely attempt to cognize it by exploiting the side-information (e.g., semantic information, etc.) you have accumulated. Inspired by this, this paper decomposes the generalized zero-shot learning (G-ZSL) task into an open set recognition (OSR) task and a zero-shot learning (ZSL) task, where OSR recognizes seen classes (if we have seen (or known) them) and rejects unseen classes (if we have never seen (or known) them before), while ZSL identifies the unseen classes rejected by the former. Simultaneously, without violating OSR's assumptions (only known class knowledge is available in training), we also first attempt to explore a new generalized open set recognition (G-OSR) by introducing the accumulated side-information from known classes to OSR. For G-ZSL, such a decomposition effectively solves the class overfitting problem with easily misclassifying unseen classes as seen classes. The problem is ubiquitous in most existing G-ZSL methods. On the other hand, for G-OSR, introducing such semantic information of known classes not only improves the recognition performance but also endows OSR with the cognitive ability of unknown classes. Specifically, a visual and semantic prototypes-jointly guided convolutional neural network (VSG-CNN) is proposed to fulfill these two tasks (G-ZSL and G-OSR) in a unified end-to-end learning framework. Extensive experiments on benchmark datasets demonstrate the advantages of our learning framework. (C) 2020 Elsevier Ltd. All rights reserved.
引用
收藏
页数:10
相关论文
共 50 条
  • [31] Open-Set Fault Recognition and Inference for Rolling Bearing Based on Open Fault Semantic Subspace
    Chen, Yu
    Tao, Laifa
    Liu, Xue
    Ma, Jian
    Lu, Chen
    Liu, Hongmei
    IEEE TRANSACTIONS ON INSTRUMENTATION AND MEASUREMENT, 2024, 73 : 1 - 11
  • [32] Leveraging Self-Distillation and Disentanglement Network to Enhance Visual-Semantic Feature Consistency in Generalized Zero-Shot Learning
    Liu, Xiaoming
    Wang, Chen
    Yang, Guan
    Wang, Chunhua
    Long, Yang
    Liu, Jie
    Zhang, Zhiyuan
    ELECTRONICS, 2024, 13 (10)
  • [33] Open-Set Recognition for Skin Lesions Using Dermoscopic Images
    Budhwant, Pranav
    Shinde, Sumeet
    Ingalhalikar, Madhura
    MACHINE LEARNING IN MEDICAL IMAGING, MLMI 2020, 2020, 12436 : 614 - 623
  • [34] Open-Set Recognition Using Intra-Class Splitting
    Schlachter, Patrick
    Liao, Yiwen
    Yang, Bin
    2019 27TH EUROPEAN SIGNAL PROCESSING CONFERENCE (EUSIPCO), 2019,
  • [35] Audio-Visual Generalized Zero-Shot Learning Based on Variational Information Bottleneck
    Li, Yapeng
    Luo, Yong
    Du, Bo
    2023 IEEE INTERNATIONAL CONFERENCE ON MULTIMEDIA AND EXPO, ICME, 2023, : 450 - 455
  • [36] Generalized zero-shot learning for action recognition with web-scale video data
    Kun Liu
    Wu Liu
    Huadong Ma
    Wenbing Huang
    Xiongxiong Dong
    World Wide Web, 2019, 22 : 807 - 824
  • [37] Generalized zero-shot learning for action recognition with web-scale video data
    Liu, Kun
    Liu, Wu
    Ma, Huadong
    Huang, Wenbing
    Dong, Xiongxiong
    WORLD WIDE WEB-INTERNET AND WEB INFORMATION SYSTEMS, 2019, 22 (02): : 807 - 824
  • [38] Enhancing Semantic-Consistent Features and Transforming Discriminative Features for Generalized Zero-Shot Classifications
    Yang, Guan
    Han, Ayou
    Liu, Xiaoming
    Liu, Yang
    Wei, Tao
    Zhang, Zhiyuan
    APPLIED SCIENCES-BASEL, 2022, 12 (24):
  • [39] Improving generalized zero-shot learning via cluster-based semantic disentangling representation
    Gao, Yi
    Feng, Wentao
    Xiao, Rong
    He, Lihuo
    He, Zhenan
    Lv, Jiancheng
    Tang, Chenwei
    PATTERN RECOGNITION, 2024, 150
  • [40] Learning Discriminative Projection With Visual Semantic Alignment for Generalized Zero Shot Learning
    Du, Pengzhen
    Zhang, Haofeng
    Lu, Jianfeng
    IEEE ACCESS, 2020, 8 (08): : 166273 - 166282