Gunfire Detection from Video
Streams audio from any video/RTSP feed and flags gunshots in under 100ms per window — no buffering delay.
How it works
A prior version of this system buffered audio and ran a heavy image-style model per clip — accurate, but too slow for a live feed. This version streams small audio windows continuously through a lightweight classifier.
Data Sourcing
Combined gunshot recordings with broad real-world sound data for realistic positive and negative coverage.
KaggleFeature Extraction
Each short audio window is reduced to a compact spectral fingerprint — not a full spectrogram image.
Librosa · MFCCModel Training
A lightweight classifier is trained and validated on a held-out split to check it generalises, not just memorises.
TensorFlow · scikit-learnReal-Time Streaming
Audio is decoded continuously from video, RTSP, or mic input in small overlapping windows — nothing is buffered.
FFmpegLive Alerting
Each window is scored in roughly 70–90ms and alerts surface instantly with a confidence score.
Python