TEST: Temporal-spatial separated transformer for temporal action localization

被引:0
|
作者
Wan, Herun [1 ,2 ,3 ]
Luo, Minnan [1 ,2 ,3 ]
Li, Zhihui [4 ]
Wang, Yang [5 ]
机构
[1] Xi An Jiao Tong Univ, Sch Comp Sci & Technol, Xian 710049, Peoples R China
[2] Xi An Jiao Tong Univ, Minist Educ, Key Lab Intelligent Networks & Network Secur, Xian 710049, Peoples R China
[3] Xi An Jiao Tong Univ, Shaanxi Prov Key Lab Big Data Knowledge Engn, Xian 710049, Peoples R China
[4] Univ Sci & Technol China, Sch Informat Sci & Technol, Hefei 230026, Peoples R China
[5] Xi An Jiao Tong Univ, Sch Continuing Educ, Xian 710049, Peoples R China
基金
中国国家自然科学基金;
关键词
Video transformer; Temporal action localization; High efficiency; NETWORK;
D O I
10.1016/j.neucom.2024.128688
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Temporal action localization is a fundamental task in video understanding. Existing methods fall into three categories: anchor-based, actionness-guided, and anchor-free. Anchor-based and actionness-guided models need huge computation resources to process redundant proposals or enumerate every possible proposal. Anchor- free models with lighter parameters become a more attractive option as temporal actions become more complex. However, they typically struggle to achieve high performance due to the need to aggregate global temporal-spatial features at every time step. To overcome this limitation, we design three efficient transformer- based architectures, bringing two advantages: (i) the global receptive field of transformers enables models to aggregate spatial and temporal at each time step, and (ii) the transformers could capture the moment-level feature, enhancing localization performance. Our designed architectures are adapted to any framework, thus we propose a simple but effective anchor-free framework named TEST. Compared to strong baselines, TEST achieves 0.96% to 3.20% improvement on two real-world datasets. Meanwhile, it improves time efficiency by 1.36 times and space efficiency by 1.08 times. Further experiments prove the effectiveness of TEST's modules. Implementation of our work is available at https://github.com/whr000001/TeST.
引用
收藏
页数:9
相关论文
共 50 条
  • [31] Spatial-temporal graph transformer network for skeleton-based temporal action segmentation
    Xiaoyan Tian
    Ye Jin
    Zhao Zhang
    Peng Liu
    Xianglong Tang
    Multimedia Tools and Applications, 2024, 83 : 44273 - 44297
  • [32] The temporal-spatial encoding of acupuncture effects in the brain
    Qin, Wei
    Bai, Lijun
    Dai, Jianping
    Liu, Peng
    Dong, Minghao
    Liu, Jixin
    Sun, Jinbo
    Yuan, Kai
    Chen, Peng
    Zhao, Baixiao
    Gong, Qiyong
    Tian, Jie
    Liu, Yijun
    MOLECULAR PAIN, 2011, 7
  • [33] Location Prediction: A Temporal-Spatial Bayesian Model
    Jia, Yantao
    Wang, Yuanzhuo
    Jin, Xiaolong
    Cheng, Xueqi
    ACM TRANSACTIONS ON INTELLIGENT SYSTEMS AND TECHNOLOGY, 2016, 7 (03)
  • [34] Spatial Enhancement and Temporal Constraint for Weakly Supervised Action Localization
    Qin, Xiaolei
    Ge, Yongxin
    Yu, Hui
    Chen, Feiyu
    Yang, Dan
    IEEE SIGNAL PROCESSING LETTERS, 2020, 27 : 1520 - 1524
  • [35] Modeling residual dynamics of helicopters based on temporal-spatial Transformer and low-rank compression
    Zhang, Hailang
    Liu, Jing
    Hu, Yu
    Tian, Xiaoqing
    NONLINEAR DYNAMICS, 2025,
  • [36] Hierarchy Spatial-Temporal Transformer for Action Recognition in Short Videos
    Cai, Guoyong
    Cai, Yumeng
    FUZZY SYSTEMS AND DATA MINING VI, 2020, 331 : 760 - 774
  • [37] 4-DIMENSIONAL TEMPORAL-SPATIAL HOLOGRAMS
    MOSSBERG
    ELECTRO-OPTICAL SYSTEMS DESIGN, 1982, 14 (03): : 52 - 52
  • [38] The Prediction of Temporal-Spatial Distribution of Space Debris
    Li, XiaoRui
    PROCEEDINGS OF THE 2016 3RD INTERNATIONAL CONFERENCE ON MATERIALS ENGINEERING, MANUFACTURING TECHNOLOGY AND CONTROL, 2016, 67 : 934 - 936
  • [39] Temporal-spatial heterogeneity of hematocrit in microvascular networks
    Li, Guansheng
    Ye, Ting
    Yang, Bo
    Wang, Sitong
    Li, Xuejin
    PHYSICS OF FLUIDS, 2023, 35 (02)
  • [40] TEMPORAL-SPATIAL CONFUSION .1. SPACE
    MARCHAIS, P
    ANNALES MEDICO-PSYCHOLOGIQUES, 1977, 135 (02): : 351 - 360