Real-Time Violence Detection via Short-Window Sliding Learning
September 4, 2026
This framework uses 1-2 second clips and LLM-based auto-captioning to train real-time violence detection models. The method achieved 95.25% accuracy on RWF-2000 and 83.25% on UCF-Crime by preserving temporal continuity through fine-grained labeling.
HOW THIS AFFECTS YOU
●
builderYou can utilize LLM auto-labeling to rapidly generate fine-grained training sets for temporal video tasks.
●
researcherThe short-window approach provides a scalable way to handle temporal continuity in surveillance datasets.