FINE-GRAINED POSE TEMPORAL MEMORY MODULE FOR VIDEO POSE ESTIMATION AND TRACKING

被引:0
|
作者
Wang, Chaoyi [1 ]
Hua, Yang [2 ]
Song, Tao [1 ]
Xue, Zhengui [1 ]
Ma, Ruhui [1 ]
Robertson, Neil [2 ]
Guan, Haibing [1 ]
机构
[1] Shanghai Jiao Tong Univ, Shanghai, Peoples R China
[2] Queens Univ Belfast, Belfast, Antrim, North Ireland
关键词
video pose estimation and tracking; keypoint occlusion;
D O I
10.1109/ICASSP39728.2021.9413650
中图分类号
O42 [声学];
学科分类号
070206 ; 082403 ;
摘要
The task of video pose estimation and tracking has been largely improved with the development of image pose estimation recently. However, there are still many challenging cases, such as body part occlusion, fast body motion, camera zooming, and complex background. Most existing methods generally use the temporal information to get more precise human bounding boxes or just use it in the tracking stage, but they fail to improve the accuracy of pose estimation tasks. To better solve these problems and utilize the temporal information efficiently and effectively, we present a novel structure, called pose temporal memory module, which is flexible to be transferred into top-down pose estimation frameworks. The temporal information stored in the pose temporal memory is aggregated into the current frame feature in our proposed module. We also transfer compositional de-attention (CoDA) to solve the unique keypoint occlusion problem in this task and propose a novel keypoint feature replacement to recover the extreme error detection under fine-grained keypoint-level guidance. To verify the generality and effectiveness of our proposed method, we integrate our module into two widely used pose estimation frameworks and obtain notable improvement on the PoseTrack dataset with only a few extra computing resources.
引用
收藏
页码:2205 / 2209
页数:5
相关论文
共 50 条
  • [31] FineAction: A Fine-Grained Video Dataset for Temporal Action Localization
    Liu, Yi
    Wang, Limin
    Wang, Yali
    Ma, Xiao
    Qiao, Yu
    IEEE TRANSACTIONS ON IMAGE PROCESSING, 2022, 31 : 6937 - 6950
  • [32] Pose Selection for Underwater Object Detection, Pose Estimation, and Tracking
    Teigland, Hakon
    Hassani, Vahid
    Tore Moller, Ments
    IEEE ACCESS, 2024, 12 : 142331 - 142342
  • [33] FinePOSE: Fine-Grained Prompt-Driven 3D Human Pose Estimation via Diffusion Models
    Xu, Jinglin
    Guo, Yijie
    Peng, Yuxin
    2024 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION, CVPR 2024, 2024, : 561 - 570
  • [34] A CRNN module for hand pose estimation
    Hu, Zhongxu
    Hu, Youmin
    Liu, Jie
    Wu, Bo
    Han, Dongmin
    Kurfess, Thomas
    NEUROCOMPUTING, 2019, 333 : 157 - 168
  • [35] FTCM: Frequency-Temporal Collaborative Module for Efficient 3D Human Pose Estimation in Video
    Tang, Zhenhua
    Hao, Yanbin
    Li, Jia
    Hong, Richang
    IEEE TRANSACTIONS ON CIRCUITS AND SYSTEMS FOR VIDEO TECHNOLOGY, 2024, 34 (02) : 911 - 923
  • [36] Robust coverless video steganography based on pose estimation and object tracking
    Li, Nan
    Qin, Jiaohua
    Xiang, Xuyu
    Tan, Yun
    JOURNAL OF INFORMATION SECURITY AND APPLICATIONS, 2024, 87
  • [37] Badminton Video Analysis Based on Player Tracking and Pose Trajectory Estimation
    Cai Cuiping
    2021 13TH INTERNATIONAL CONFERENCE ON MEASURING TECHNOLOGY AND MECHATRONICS AUTOMATION (ICMTMA 2021), 2021, : 471 - 474
  • [38] Bidirectional Temporal Pose Matching for Tracking
    Fang, Yichuan
    Shi, Qingxuan
    Yang, Zhen
    ELECTRONICS, 2024, 13 (02)
  • [39] EHTracker: Toward Fine-Grained Localization for Satellite Video Target Tracking
    Yang, Jianwei
    Liu, Yuhan
    Liu, Yanxing
    Wang, Ziming
    Li, Jiawei
    Zhou, Guangyao
    Wang, Wenzhi
    Hu, Yuxin
    IEEE JOURNAL OF SELECTED TOPICS IN APPLIED EARTH OBSERVATIONS AND REMOTE SENSING, 2025, 18 : 5582 - 5599
  • [40] Hands-on: deformable pose and motion models for spatiotemporal localization of fine-grained dyadic interactions
    Coert van Gemeren
    Ronald Poppe
    Remco C. Veltkamp
    EURASIP Journal on Image and Video Processing, 2018