Guided CNN for generalized zero-shot and open-set recognition using visual and semantic prototypes

被引：32

作者：

Geng, Chuanxing ^{[1
]}

Tao, Lue ^{[1
]}

Chen, Songcan ^{[1
]}

机构：

[1] Nanjing Univ Aeronaut & Astronaut, Coll Comp Sci & Technol, MIIT Key Lab Pattern Anal & Machine Intelligence, Nanjing 211106, Peoples R China

来源：

PATTERN RECOGNITION | 2020年 / 102卷

关键词：

Convolutional prototype learning; Generalized zero-shot Learning; Open set recognition;

D O I：

10.1016/j.patcog.2020.107263

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

In the process of exploring the world, the curiosity constantly drives humans to cognize new things. Supposing you are a zoologist, for a presented animal image, you can recognize it immediately if you know its class. Otherwise, you would more likely attempt to cognize it by exploiting the side-information (e.g., semantic information, etc.) you have accumulated. Inspired by this, this paper decomposes the generalized zero-shot learning (G-ZSL) task into an open set recognition (OSR) task and a zero-shot learning (ZSL) task, where OSR recognizes seen classes (if we have seen (or known) them) and rejects unseen classes (if we have never seen (or known) them before), while ZSL identifies the unseen classes rejected by the former. Simultaneously, without violating OSR's assumptions (only known class knowledge is available in training), we also first attempt to explore a new generalized open set recognition (G-OSR) by introducing the accumulated side-information from known classes to OSR. For G-ZSL, such a decomposition effectively solves the class overfitting problem with easily misclassifying unseen classes as seen classes. The problem is ubiquitous in most existing G-ZSL methods. On the other hand, for G-OSR, introducing such semantic information of known classes not only improves the recognition performance but also endows OSR with the cognitive ability of unknown classes. Specifically, a visual and semantic prototypes-jointly guided convolutional neural network (VSG-CNN) is proposed to fulfill these two tasks (G-ZSL and G-OSR) in a unified end-to-end learning framework. Extensive experiments on benchmark datasets demonstrate the advantages of our learning framework. (C) 2020 Elsevier Ltd. All rights reserved.

引用

页数：10

共 50 条

[1] Indirect visual–semantic alignment for generalized zero-shot recognition
Yan-He Chen
Mei-Chen Yeh
Multimedia Systems, 2024, 30
[2] Vocabulary-Informed Zero-Shot and Open-Set Learning
Fu, Yanwei
Wang, Xiaomei
Dong, Hanze
Jiang, Yu-Gang
Wang, Meng
Xue, Xiangyang
Sigal, Leonid
IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2020, 42 (12) : 3136 - 3152
[3] Indirect visual-semantic alignment for generalized zero-shot recognition
Chen, Yan-He
Yeh, Mei-Chen
MULTIMEDIA SYSTEMS, 2024, 30 (02)
[4] CoHOZ: Contrastive Multimodal Prompt Tuning for Hierarchical Open-set Zero-shot Recognition
Liao, Ning
Liu, Yifeng
Li, Xiaobo
Lei, Chenyi
Wang, Guoxin
Hua, Xian-Sheng
Yan, Junchi
PROCEEDINGS OF THE 30TH ACM INTERNATIONAL CONFERENCE ON MULTIMEDIA, MM 2022, 2022, : 3262 - 3271
[5] Learning multiple gaussian prototypes for open-set recognition
Liu, Jiaming
Tian, Jun
Han, Wei
Qin, Zhili
Fan, Yulu
Shao, Junming
INFORMATION SCIENCES, 2023, 626 : 738 - 753
[6] Feature-semantic augmentation network for few-shot open-set recognition
Huang, Xilang
Choi, Seon Han
PATTERN RECOGNITION, 2024, 156
[7] Semantic Contrastive Embedding for Generalized Zero-Shot Learning
Han, Zongyan
Fu, Zhenyong
Chen, Shuo
Yang, Jian
INTERNATIONAL JOURNAL OF COMPUTER VISION, 2022, 130 (11) : 2606 - 2622
[8] Semantic Contrastive Embedding for Generalized Zero-Shot Learning
Zongyan Han
Zhenyong Fu
Shuo Chen
Jian Yang
International Journal of Computer Vision, 2022, 130 : 2606 - 2622
[9] Few-shot Open-set Recognition Using Background as Unknowns
Song, Nan
Zhang, Chi
Lin, Guosheng
PROCEEDINGS OF THE 30TH ACM INTERNATIONAL CONFERENCE ON MULTIMEDIA, MM 2022, 2022, : 5970 - 5979
[10] Joint Visual and Semantic Optimization for zero-shot learning
Wu, Hanrui
Yan, Yuguang
Chen, Sentao
Huang, Xiangkang
Wu, Qingyao
Ng, Michael K.
KNOWLEDGE-BASED SYSTEMS, 2021, 215 (215)

← 1 2 3 4 5 →