PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
April 4, 202119 citationsOpen Access

MIST: Multiple Instance Self-Training Framework for Video Anomaly Detection

JFJia-Chang FengFHFa-Ting HongWZWei‐Shi Zheng

Key Points

Key points are not available for this paper at this time.

Abstract

Weakly supervised video anomaly detection (WS-VAD) is to distinguish anomalies from normal events based on discriminative representations. Most existing works are limited in insufficient video representations. In this work, we develop a multiple instance self-training framework (MIST)to efficiently refine task-specific discriminative representations with only video-level annotations. In particular, MIST is composed of 1) a multiple instance pseudo label generator, which adapts a sparse continuous sampling strategy to produce more reliable clip-level pseudo labels, and 2) a self-guided attention boosted feature encoder that aims to automatically focus on anomalous regions in frames while extracting task-specific representations. Moreover, we adopt a self-training scheme to optimize both components and finally obtain a task-specific feature encoder. Extensive experiments on two public datasets demonstrate the efficacy of our method, and our method performs comparably to or even better than existing supervised and weakly supervised methods, specifically obtaining a frame-level AUC 94.83% on ShanghaiTech.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Feng et al. (2021) studied this question.

synapsesocial.com/papers/6a11c11f224d77b8d5617a4ehttps://doi.org/10.48550/arxiv.2104.01633
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Deep Learning is Robust to Massive Label Noise2017 · 406 citations
  2. 2Unsupervised Domain Adaptation for Semantic Segmentation via Class-Balanced Self-training2018 · 1,502 citations
  3. 3Weakly Supervised Coupled Networks for Visual Sentiment Analysis2018 · 167 citations
  4. 4The Kinetics Human Action Video Dataset2017 · 2,894 citations
  5. 5CNN features with bi-directional LSTM for real-time anomaly detection in surveillance networks2020 · 308 citations