TY - JOUR
T1 - Robust video anomaly detection via causal feature-guided data augmentation
AU - Li, Zhaoyi
AU - Shi, Chenrui
AU - Sun, Che
AU - Wu, Yuwei
N1 - Publisher Copyright:
© 2026 Elsevier Inc.
PY - 2026/3
Y1 - 2026/3
N2 - Data augmentation is a commonly used technique to learn distinguishing features for accurately identifying video anomaly patterns. However, traditional data augmentation methods introduce spurious features-spurious correlations (e.g. , lighting conditions) that lack causal links to anomalies, yielding models that excel on specific test sets but fail in real-world scenarios with varying conditions. In this paper, we propose a novel robust video anomaly detection method by designing a causal feature-guided augmentation technique that strategically enhances the learnability of predictive patterns while suppressing spurious correlations. Specifically, we first disentangle causal features , i.e. , directly predictive of labels, from spurious features via a causal generative model. We then perturb causal features to enhance their variability in training data, compelling the model to focus on invariant patterns while ignoring spurious correlations. The augmented samples are reconstructed with the refined causal representations, enhancing the model’s discriminative capability. Furthermore, we introduce a systematic evaluation framework with three increasing difficulty levels to assess robustness: seen/unseen variations, cross-dataset generalization, and cross-domain adaptation. This evaluation examines system stability under varying conditions, closely aligning with real-world surveillance deployment requirements. Experimental results validate both the effectiveness and robustness of our method.
AB - Data augmentation is a commonly used technique to learn distinguishing features for accurately identifying video anomaly patterns. However, traditional data augmentation methods introduce spurious features-spurious correlations (e.g. , lighting conditions) that lack causal links to anomalies, yielding models that excel on specific test sets but fail in real-world scenarios with varying conditions. In this paper, we propose a novel robust video anomaly detection method by designing a causal feature-guided augmentation technique that strategically enhances the learnability of predictive patterns while suppressing spurious correlations. Specifically, we first disentangle causal features , i.e. , directly predictive of labels, from spurious features via a causal generative model. We then perturb causal features to enhance their variability in training data, compelling the model to focus on invariant patterns while ignoring spurious correlations. The augmented samples are reconstructed with the refined causal representations, enhancing the model’s discriminative capability. Furthermore, we introduce a systematic evaluation framework with three increasing difficulty levels to assess robustness: seen/unseen variations, cross-dataset generalization, and cross-domain adaptation. This evaluation examines system stability under varying conditions, closely aligning with real-world surveillance deployment requirements. Experimental results validate both the effectiveness and robustness of our method.
KW - Causal generative model
KW - Data augmentation
KW - Robustness analysis
KW - Video anomaly detection
UR - https://www.scopus.com/pages/publications/105028984817
U2 - 10.1016/j.cviu.2026.104671
DO - 10.1016/j.cviu.2026.104671
M3 - Article
AN - SCOPUS:105028984817
SN - 1077-3142
VL - 265
JO - Computer Vision and Image Understanding
JF - Computer Vision and Image Understanding
M1 - 104671
ER -