← Back to literature explorer
Unsupervised & Semi-Supervised From survey PDF — verify links

SFN-VAD

SFN-VAD — Prediction + Semantic Consistency

arXiv · 2025

Survey summary

Fuses appearance, motion, and language-derived semantic cues in a tri-modal encoder; a multimodal decoder predicts next frame and semantic representation, and a Sparse Feature Filtering Module limits over-generalization to abnormal events.

Unsupervised & Semi-Supervised

Unsupervised & Semi-SupervisedPrediction + Semantic Consistency
Unsupervised & Semi-SupervisedTri-modalSparse filteringPrediction error