Learning Scene-Aware Spatio-Temporal GNNs for Few-Shot Early Action Prediction

被引:8
作者
Hu, Yufan [1 ,2 ]
Gao, Junyu [3 ,4 ]
Xu, Changsheng [2 ,3 ,4 ]
机构
[1] Hefei Univ Technol, Hefei 230009, Peoples R China
[2] Peng Cheng Lab, Shenzhen 518055, Peoples R China
[3] Chinese Acad Sci, Inst Automat, Natl Lab Pattern Recognit, Beijing 100190, Peoples R China
[4] Univ Chinese Acad Sci, Sch Artificial Intelligence, Natl Lab Pattern Recognit, Beijing 100190, Peoples R China
基金
北京市自然科学基金; 中国国家自然科学基金;
关键词
Few-shot learning; early action prediction; scene graph; graph neural network; OBJECT AFFORDANCES;
D O I
10.1109/TMM.2022.3142413
中图分类号
TP [自动化技术、计算机技术];
学科分类号
0812 ;
摘要
We aim to address a new task named few-shot early action prediction (FS-EAP) that learns classifiers for novel actions from only a few partially observed videos. We argue that the task is extremely challenging since the partially observed videos do not contain enough action information in a few-shot environment. To tackle this task, in this paper, we propose a scene-aware spatio-temporal graph neural network (SA-STGNN) by leveraging the fine-grained spatio-temporal interactions in the video scenes. Specifically, we first generate a spatio-temporal graph corresponding to the partially observed video to capture comprehensive spatio-temporal correlations. Then we utilize the spatio-temporal graph as the input of our SA-STGNN and predict the augmented video features corresponding to the complete video. The architecture uses several scene-aware learning blocks, which are a combination of edge fusion graph neural layers and temporal gated convolutional layers to jointly model spatial and temporal dependencies. Finally, we employ an early action predictor to exploit the learned video features for predicting actions in the few-shot setting. Extensive experimental results on two widely adopted video datasets demonstrate the effectiveness of our approach and its superior performance over the state-of-the-art approaches.
引用
收藏
页码:2061 / 2073
页数:13
相关论文
共 50 条
  • [21] Personalized Multiparty Few-Shot Learning for Remote Sensing Scene Classification
    Wang, Shanfeng
    Li, Jianzhao
    Liu, Zaitian
    Gong, Maoguo
    Zhang, Yourun
    Zhao, Yue
    Deng, Boya
    Zhou, Yu
    IEEE TRANSACTIONS ON GEOSCIENCE AND REMOTE SENSING, 2024, 62 : 1 - 15
  • [22] Dynamic Relation-Aware Multiple Instance Learning for Few-shot Learning
    Zheng, Kaipeng
    Cheng, Liu
    Shen, Jie
    2022 INTERNATIONAL JOINT CONFERENCE ON NEURAL NETWORKS (IJCNN), 2022,
  • [23] MetaHKG: Meta Hyperbolic Learning for Few-shot Temporal Reasoning
    Wang, Ruijie
    Zhang, Yutong
    Li, Jinyang
    Liu, Shengzhong
    Sun, Dachun
    Wang, Tianchen
    Wang, Tianshi
    Chen, Yizhuo
    Kara, Denizhan
    Abdelzaher, Tarek
    PROCEEDINGS OF THE 47TH INTERNATIONAL ACM SIGIR CONFERENCE ON RESEARCH AND DEVELOPMENT IN INFORMATION RETRIEVAL, SIGIR 2024, 2024, : 59 - 69
  • [24] VISUAL TEMPO CONTRASTIVE LEARNING FOR FEW-SHOT ACTION RECOGNITION
    Wang, Guangge
    Ye, Weirong
    Wang, Xiao
    Jin, Rongrong
    Wang, Hanzi
    2022 IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING, ICIP, 2022, : 1096 - 1100
  • [25] Few-Shot Action Recognition with Hierarchical Matching and Contrastive Learning
    Zheng, Sipeng
    Chen, Shizhe
    Jin, Qin
    COMPUTER VISION - ECCV 2022, PT IV, 2022, 13664 : 297 - 313
  • [26] Task-aware prototype refinement for improved few-shot learning
    Wei Zhang
    Xiaodong Gu
    Neural Computing and Applications, 2023, 35 : 17899 - 17913
  • [27] Mutually-aware feature learning for few-shot object counting
    Jeon, Yerim
    Lee, Subeen
    Kim, Jihwan
    Heo, Jae-Pil
    PATTERN RECOGNITION, 2025, 161
  • [28] A Novel Group-Aware Pruning Method for Few-shot Learning
    Zheng, Yin-Dong
    Ma, Yun-Tao
    Liu, Ruo-Ze
    Lu, Tong
    2019 INTERNATIONAL JOINT CONFERENCE ON NEURAL NETWORKS (IJCNN), 2019,
  • [29] Hierarchy-Aware Interactive Prompt Learning for Few-Shot Classification
    Yin, Xiaotian
    Wu, Jiamin
    Yang, Wenfei
    Zhou, Xu
    Zhang, Shifeng
    Zhang, Tianzhu
    IEEE TRANSACTIONS ON CIRCUITS AND SYSTEMS FOR VIDEO TECHNOLOGY, 2024, 34 (12) : 12221 - 12232
  • [30] Bidirectional Patch-Aware Attention Network for Few-Shot Learning
    Mao, Yu
    Lin, Shaojie
    Lin, Zilong
    Lin, Yaojin
    IEEE TRANSACTIONS ON COMPUTATIONAL SOCIAL SYSTEMS, 2025,