Self-Training Multi-Sequence Learning with Transformer for Weakly Supervised Video Anomaly Detection

Li, Shuo; Liu, Fang; Jiao, Licheng

doi:10.1609/aaai.v36i2.20028

articleProceedings of the AAAI Conference on Artificial IntelligenceJun 28, 2022DIAMOND OA

Self-Training Multi-Sequence Learning with Transformer for Weakly Supervised Video Anomaly Detection

SLShuo Li FLFang Liu LJLicheng Jiao

Xidian University

Indexed incrossref

Abstract

Weakly supervised Video Anomaly Detection (VAD) using Multi-Instance Learning (MIL) is usually based on the fact that the anomaly score of an abnormal snippet is higher than that of a normal snippet. In the beginning of training, due to the limited accuracy of the model, it is easy to select the wrong abnormal snippet. In order to reduce the probability of selection errors, we first propose a Multi-Sequence Learning (MSL) method and a hinge-based MSL ranking loss that uses a sequence composed of multiple snippets as an optimization unit. We then design a Transformer-based MSL network to learn both video-level anomaly probability and snippet-level anomaly scores. In the inference stage, we propose to use the…

Citation impact

236

total citations

FWCI: 22.54
Percentile: 100%
References: 62

Citations per year

Authors

3

Topics & keywords

Topics

Keywords

Snippet
Anomaly detection
Computer science
Anomaly (physics)
Inference
Artificial intelligence
Sequence (biology)
Transformer

UN Sustainable Development Goals

Peace, Justice and strong institutions

No related works found for this paper.