MARATTO

article

An Attention-Based Unsupervised Framework for Video Anomaly Detection in Large Heterogeneous Environments

20241 citationBenha University

Abstract

Although automated detection capabilities are frequently included in modern video anomaly detection systems, these algorithms may still encounter issues including high false-positive rates and trouble comprehending complicated circumstances. Dealing with big heterogeneous datasets that are taken in various lighting and environmental circumstances with changing resolution and quality results in obstacles for video anomaly detection systems. This paper proposes an end-to-end vision transformer (ViT)-based framework for video anomaly detection in surveillance systems. The proposed framework incorporates the ViT model with a spatio-temporal attention model augmented with Long Short-Term Memory (LSTM) layers to extract both the global context and the spatio-temporal features of the video frames. The framework is mainly designed to work with large and heterogeneous datasets that represent different environments, which remains challenging for the current video anomaly detection efforts. To assess the proposed framework, video anomaly benchmark datasets are used: the big ShanghaiTech dataset, the CUHK Avenue dataset, and the VCSD-Ped2 dataset. The proposed framework achieved AUC_ROC of 79.6%, 83.4%, and 93.8% for the three datasets, respectively. This illustrates the superior performance of the proposed framework in identifying video abnormalities compared to SOTA, especially for large heterogeneous datasets, i.e., ShanghaiTech.

Research topics

  • Anomaly Detection Techniques and Applications
  • Network Security and Intrusion Detection
  • Digital Media Forensic Detection

Sustainable Development Goals

Read the original research

This page summarises published work. The authoritative version sits with the publisher.

DOI: 10.1109/miucc62295.2024.10783549

Is something wrong with this record? Report it or request removal.

Discussion

Discuss this research

Have you built on this work, tried to replicate it, or seen it applied in practice? Share what you know. Verified researchers and MARATTO™ domain experts can open a discussion, and any member can reply. Contributions are reviewed before they appear.

No discussion yet. Open the first thread.