跳到主要导航 跳到搜索 跳到主要内容

Multi-frame feature-fusion-based model for violence detection

  • Shanghai Jiao Tong University
  • University of Technology Sydney

科研成果: 期刊稿件文章同行评审

摘要

Human behavior detection is essential for public safety and monitoring. However, in human-based surveillance systems, it requires continuous human attention and observation, which is a difficult task. Detection of violent human behavior using autonomous surveillance systems is of critical importance for uninterrupted video surveillance. In this paper, we propose a novel method to detect fights or violent actions based on learning both the spatial and temporal features from equally spaced sequential frames of a video. Multi-level features for two sequential frames, extracted from the convolutional neural network’s top and bottom layers, are combined using the proposed feature fusion method to take into account the motion information. We also proposed Wide-Dense Residual Block to learn these combined spatial features from the two input frames. These learned features are then concatenated and fed to long short-term memory units for capturing temporal dependencies. The feature fusion method and use of additional wide-dense residual blocks enable the network to learn combined features from the input frames effectively and yields better accuracy results. Experimental results evaluated on four publicly available datasets: HockeyFight, Movies, ViolentFlow and BEHAVE show the superior performance of the proposed model in comparison with the state-of-the-art methods.

源语言英语
页(从-至)1415-1431
页数17
期刊Visual Computer
37
6
DOI
出版状态已出版 - 6月 2021
已对外发布

联合国可持续发展目标

此成果有助于实现下列可持续发展目标:

  1. 可持续发展目标 16 - 和平、正义和强大机构
    可持续发展目标 16 和平、正义和强大机构

学术指纹

探究 'Multi-frame feature-fusion-based model for violence detection' 的科研主题。它们共同构成独一无二的学术指纹。

引用此