VideoClusterNet: Self-supervised and Adaptive Face Clustering for Videos

被引:0
|
作者
Walawalkar, Devesh [1 ]
Garrido, Pablo [1 ]
机构
[1] Flawless AI, London, England
来源
关键词
REPRESENTATION;
D O I
10.1007/978-3-031-73404-5_22
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
With the rise of digital media content production, the need for analyzing movies and TV series episodes to locate the main cast of characters precisely is gaining importance. Specifically, Video Face Clustering aims to group together detected video face tracks with common facial identities. This problem is very challenging due to the large range of pose, expression, appearance, and lighting variations of a given face across video frames. Generic pre-trained Face Identification (ID) models fail to adapt well to the video production domain, given its high dynamic range content and also unique cinematic style. Furthermore, traditional clustering algorithms depend on hyperparameters requiring individual tuning across datasets. In this paper, we present a novel video face clustering approach that learns to adapt a generic face ID model to new video face tracks in a fully self-supervised fashion. We also propose a parameter-free clustering algorithm that is capable of automatically adapting to the finetuned model's embedding space for any input video. Due to the lack of comprehensive movie face clustering benchmarks, we also present a first-of-kind movie dataset: MovieFaceCluster. Our dataset is handpicked by film industry professionals and contains extremely challenging face ID scenarios. Experiments show our method's effectiveness in handling difficult mainstream movie scenes on our benchmark dataset and state-of-the-art performance on traditional TV series datasets.
引用
收藏
页码:377 / 396
页数:20
相关论文
共 50 条
  • [31] StepFormer: Self-supervised Step Discovery and Localization in Instructional Videos
    Dvornik, Nikita
    Hadji, Isma
    Zhang, Ran
    Derpanis, Konstantinos G.
    Wildes, Richard P.
    Jepson, Allan D.
    2023 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2023, : 18952 - 18961
  • [32] Segmenting Cardiac Ultrasound Videos Using Self-Supervised Learning
    Lamoureux, Erik
    Ayromlou, Sana
    Amiri, Seyedeh Neda Ahmadi
    Rhodin, Helge
    2023 45TH ANNUAL INTERNATIONAL CONFERENCE OF THE IEEE ENGINEERING IN MEDICINE & BIOLOGY SOCIETY, EMBC, 2023,
  • [33] Self-Supervised Object Detection and Retrieval Using Unlabeled Videos
    Amrani, Elad
    Ben-Ari, Rami
    Shapira, Inbar
    Hakim, Tal
    Bronstein, Alex
    2020 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION WORKSHOPS (CVPRW 2020), 2020, : 4100 - 4108
  • [34] Self-Supervised Human Depth Estimation from Monocular Videos
    Tan, Feitong
    Zhu, Hao
    Cui, Zhaopeng
    Zhu, Siyu
    Pollefeys, Marc
    Tan, Ping
    2020 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2020, : 647 - 656
  • [35] Semi-supervised learning made simple with self-supervised clustering
    Fini, Enrico
    Astolfi, Pietro
    Alahari, Karteek
    Alameda-Meda, Xavier
    Mairal, Julien
    Nabi, Moin
    Ricci, Elisa
    2023 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION, CVPR, 2023, : 3187 - 3197
  • [36] Adaptive Self-Supervised Graph Representation Learning
    Gong, Yunchi
    36TH INTERNATIONAL CONFERENCE ON INFORMATION NETWORKING (ICOIN 2022), 2022, : 254 - 259
  • [37] Self-supervised Adaptive Aggregator Learning on Graph
    Lin, Bei
    Luo, Binli
    He, Jiaojiao
    Gui, Ning
    ADVANCES IN KNOWLEDGE DISCOVERY AND DATA MINING, PAKDD 2021, PT III, 2021, 12714 : 29 - 41
  • [38] Adaptive self-supervised learning for sequential recommendation
    Sun, Xiujuan
    Sun, Fuzhen
    Zhang, Zhiwei
    Li, Pengcheng
    Wang, Shaoqing
    NEURAL NETWORKS, 2024, 179
  • [39] Self-supervised learning for clustering of wireless spectrum activity
    Milosheski, Ljupcho
    Cerar, Gregor
    Bertalanic, Blaz
    Fortuna, Carolina
    Mohorcic, Mihael
    COMPUTER COMMUNICATIONS, 2023, 212 : 353 - 365
  • [40] Deep Self-Supervised Hierarchical Clustering for Speaker Diarization
    Singh, Prachi
    Ganapathy, Sriram
    INTERSPEECH 2020, 2020, : 294 - 298