VideoClusterNet: Self-supervised and Adaptive Face Clustering for Videos

被引:0
|
作者
Walawalkar, Devesh [1 ]
Garrido, Pablo [1 ]
机构
[1] Flawless AI, London, England
来源
COMPUTER VISION - ECCV 2024, PT XXX | 2025年 / 15088卷
关键词
REPRESENTATION;
D O I
10.1007/978-3-031-73404-5_22
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
With the rise of digital media content production, the need for analyzing movies and TV series episodes to locate the main cast of characters precisely is gaining importance. Specifically, Video Face Clustering aims to group together detected video face tracks with common facial identities. This problem is very challenging due to the large range of pose, expression, appearance, and lighting variations of a given face across video frames. Generic pre-trained Face Identification (ID) models fail to adapt well to the video production domain, given its high dynamic range content and also unique cinematic style. Furthermore, traditional clustering algorithms depend on hyperparameters requiring individual tuning across datasets. In this paper, we present a novel video face clustering approach that learns to adapt a generic face ID model to new video face tracks in a fully self-supervised fashion. We also propose a parameter-free clustering algorithm that is capable of automatically adapting to the finetuned model's embedding space for any input video. Due to the lack of comprehensive movie face clustering benchmarks, we also present a first-of-kind movie dataset: MovieFaceCluster. Our dataset is handpicked by film industry professionals and contains extremely challenging face ID scenarios. Experiments show our method's effectiveness in handling difficult mainstream movie scenes on our benchmark dataset and state-of-the-art performance on traditional TV series datasets.
引用
收藏
页码:377 / 396
页数:20
相关论文
共 50 条
  • [31] Similarity contrastive estimation for image and video soft contrastive self-supervised learning
    Julien Denize
    Jaonary Rabarisoa
    Astrid Orcesi
    Romain Hérault
    Machine Vision and Applications, 2023, 34
  • [32] Traffic Accident Detection via Self-Supervised Consistency Learning in Driving Scenarios
    Fang, Jianwu
    Qiao, Jiahuan
    Bai, Jie
    Yu, Hongkai
    Xue, Jianru
    IEEE TRANSACTIONS ON INTELLIGENT TRANSPORTATION SYSTEMS, 2022, 23 (07) : 9601 - 9614
  • [33] Identification of extracellular vesicles from their Raman spectra via self-supervised learning
    Jensen, Mathias N.
    Guerreiro, Eduarda M.
    Enciso-Martinez, Agustin
    Kruglik, Sergei G.
    Otto, Cees
    Snir, Omri
    Ricaud, Benjamin
    Helleso, Olav Gaute
    SCIENTIFIC REPORTS, 2024, 14 (01):
  • [34] Self-Supervised Feature Learning Based on Spectral Masking for Hyperspectral Image Classification
    Liu, Weiwei
    Liu, Kai
    Sun, Weiwei
    Yang, Gang
    Ren, Kai
    Meng, Xiangchao
    Peng, Jiangtao
    IEEE TRANSACTIONS ON GEOSCIENCE AND REMOTE SENSING, 2023, 61
  • [35] Investigation of Ensemble features of Self-Supervised Pretrained Models for Automatic Speech Recognition
    Arunkumar, A.
    Sukhadia, Vrunda Nileshkumar
    Umesh, Srinivasan
    INTERSPEECH 2022, 2022, : 5145 - 5149
  • [36] Self-supervised facial expression recognition with fine-grained feature selection
    An, Heng-Yu
    Jia, Rui-Sheng
    VISUAL COMPUTER, 2024, 40 (10): : 7001 - 7013
  • [37] Self-supervised learning of grasp dependent tool affordances on the iCub Humanoid robot
    Mar, Tanis
    Tikhanoff, Vadim
    Metta, Giorgio
    Natale, Lorenzo
    2015 IEEE INTERNATIONAL CONFERENCE ON ROBOTICS AND AUTOMATION (ICRA), 2015, : 3200 - 3206
  • [38] Contrastive self-supervised learning: review, progress, challenges and future research directions
    Kumar, Pranjal
    Rawat, Piyush
    Chauhan, Siddhartha
    INTERNATIONAL JOURNAL OF MULTIMEDIA INFORMATION RETRIEVAL, 2022, 11 (04) : 461 - 488
  • [39] A multi-resolution self-supervised learning framework for semantic segmentation in histopathology
    Wang, Hao
    Ahn, Euijoon
    Kim, Jinman
    PATTERN RECOGNITION, 2024, 155
  • [40] Learnable Layer Selection and Model Fusion for Speech Self-Supervised Learning Models
    Chiu, Sheng-Chieh
    Wu, Chia-Hua
    Hsieh, Jih-Kang
    Tsao, Yu
    Wang, Hsin-Min
    INTERSPEECH 2024, 2024, : 3914 - 3918