← Back to literature explorer
Weakly Supervised From survey PDF — verify links

DWFF-VAD

DWFF-VAD — Multimodal Fusion

ICCE-Asia · 2024

Survey summary

Weakly supervised VAD using CLIP to align video and text representations, with dynamic weighting that emphasizes abnormal frames during training and a fusion of video-level and frame-level features to address the mix of normal and abnormal content within anomalous videos.

Weakly Supervised

Weakly SupervisedMultimodal Fusion
Weakly SupervisedDynamic-weighted fusion