Ravi-Teja-konda/Surveillance_Video_Summarizer
VLM driven tool that processes surveillance videos, extracts frames, and generates insightful annotations using a fine-tuned Florence-2 Vision-Language Model. Includes a Gradio-based interface for querying and analyzing video footage.
⭐ 134
⑂ 18
Python
· 2025-06-07推送
134
Watchers
0
贡献者
0
Commits
0
Releases
1
Open Issues
2025-06-07
最近推送
原文
中文
暂无 README