article Open access

GCN-Transformer Autoencoder with Knowledge Distillation for Unsupervised Video Anomaly Detection

  • Journal of Advanced Computational Intelligence and Intelligent Informatics
  • Fuji Technology Press Ltd.
Research footprint

At a glance

Citations
0
References
36
Comments
0
Paper overview

Abstract

Video anomaly detection is crucial in intelligent surveillance, yet the scarcity and diversity of abnormal events pose significant challenges for supervised methods. This paper presents an unsupervised framework that integrates graph attention networks (GATs) and Transformer architectures, combining masked autoencoders (MAEs) with self-distillation training. GATs are utilized to model spatial and inter-frame relationships, while Transformers capture long-range temporal dependencies, overcoming the limitations of traditional MAE and self-distillation approaches. The model employs a two-stage training process: first, a lightweight MAE combined with a GAT-Transformer fusion constructs a knowledge distillation module; second, the student autoencoder is optimized by integrating a graph convolutional autoencoder and a classification head to identify synthetic anomalies. We evaluate the proposed method on three representative datasets—ShanghaiTech Campus, UBnormal, and UCSD Ped2—and achieve promising results.

Record transparency

Publication details

DOI
10.20965/jaciii.2025.p0659
OpenAlex
W4410495708
Document type
article
Language
EN
Source
Journal of Advanced Computational Intelligence and Intelligent Informatics
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.