自编码
计算机科学
异常检测
编码器
人工智能
一般化
特征(语言学)
模式识别(心理学)
深度学习
数学分析
语言学
哲学
数学
操作系统
作者
Yan Fu,Bao Jian Yang,Ou Ye
出处
期刊:Electronics
[Multidisciplinary Digital Publishing Institute]
日期:2024-01-14
卷期号:13 (2): 353-353
被引量:12
标识
DOI:10.3390/electronics13020353
摘要
Video anomaly detection is a critical component of intelligent video surveillance systems, extensively deployed and researched in industry and academia. However, existing methods have a strong generalization ability for predicting anomaly samples. They cannot utilize high-level semantic and temporal contextual information in videos, resulting in unstable prediction performance. To alleviate this issue, we propose an encoder–decoder model named SMAMS, based on spatiotemporal masked autoencoder and memory modules. First, we represent and mask some of the video events using spatiotemporal cubes. Then, the unmasked patches are inputted into the spatiotemporal masked autoencoder to extract high-level semantic and spatiotemporal features of the video events. Next, we add multiple memory modules to store unmasked video patches of different feature layers. Finally, skip connections are introduced to compensate for crucial information loss caused by the memory modules. Experimental results show that the proposed method outperforms state-of-the-art methods, achieving AUC scores of 99.9%, 94.8%, and 78.9% on the UCSD Ped2, CUHK Avenue, and Shanghai Tech datasets.
科研通智能强力驱动
Strongly Powered by AbleSci AI