A temporal fusion approach for video classification with convolutional and LSTM neural networks applied to violence detection

Jean Phelipe De Oliveira Lima; Carlos Maurício Seródio Figueiredo

Journal ArticleOPEN ACCESS

A temporal fusion approach for video classification with convolutional and LSTM neural networks applied to violence detection

Inteligencia Artificial (2021) 24(67) 40-50

DOI: 10.4114/intartif.vol24iss67pp40-50

24Citations

34Readers

Abstract

In modern smart cities, there is a quest for the highest level of integration and automation service. In the surveillance sector, one of the main challenges is to automate the analysis of videos in real-time to identify critical situations. This paper presents intelligent models based on Convolutional Neural Networks (in which the MobileNet, InceptionV3 and VGG16 networks had used), LSTM networks and feedforward networks for the task of classifying videos under the classes “Violence” and “Non-Violence”, using for this the RLVS database. Different data representations held used according to the Temporal Fusion techniques. The best outcome achieved was 0.91 and 0.90 of Accuracy and F1-Score, respectively, a higher result compared to those found in similar researches for works conducted on the same database.

Author supplied keywords

Cite

CITATION STYLE

APA

De Oliveira Lima, J. P., & Figueiredo, C. M. S. (2021). A temporal fusion approach for video classification with convolutional and LSTM neural networks applied to violence detection. Inteligencia Artificial, 24(67), 40–50. https://doi.org/10.4114/intartif.vol24iss67pp40-50

A temporal fusion approach for video classification with convolutional and LSTM neural networks applied to violence detection

Abstract

Author supplied keywords

Cite

Register to see more suggestions