arXiv · 2412.02127
Streamlining Video Analysis for Efficient Violence Detection
Abstract
This paper addresses the challenge of automated violence detection in video frames captured by surveillance cameras, specifically focusing on classifying scenes as "fight" or "non-fight." This task is critical for enhancing unmanned security systems, online content filtering, and related applications. We propose an approach using a 3D Convolutional Neural Network (3D CNN)-based model named X3D to tackle this problem. Our approach incorporates pre-processing steps such as tube extraction, volume cropping, and frame aggregation, combined with clustering techniques, to accurately localize and classify fight scenes. Extensive experimentation demonstrates the effectiveness of our method in distinguishing violent from non-violent events, providing valuable insights for advancing practical violence detection systems.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Gourang Pathak, Abhay Kumar, Sannidhya Rawat, Shikha Gupta. 2024-11-29. Streamlining Video Analysis for Efficient Violence Detection. https://arxiv.org/abs/2412.02127
Cite the original work for its findings. Save a collection to share your selection of sources.