Weakly Supervised From survey PDF — verify links
DWFF-VAD
DWFF-VAD — Multimodal Fusion
ICCE-Asia · 2024Human-reviewed summary status
Survey summary
Weakly supervised VAD using CLIP to align video and text representations, with dynamic weighting that emphasizes abnormal frames during training and a fusion of video-level and frame-level features to address the mix of normal and abnormal content within anomalous videos.
Taxonomy placement
Weakly Supervised
Weakly SupervisedMultimodal Fusion
Method characteristics
Weakly SupervisedDynamic-weighted fusion