{"repo":"arashsajjadi/ai-powered-video-analyzer","free":true,"listed":false,"github":"https://github.com/arashsajjadi/ai-powered-video-analyzer","clone":"git clone https://github.com/arashsajjadi/ai-powered-video-analyzer.git","description":"An offline AI-powered video analysis tool with object detection (YOLO), image captioning (BLIP), speech transcription (Whisper), audio event detection (PANNs), and AI-generated summaries (LLMs via Ollama). It ensures privacy and offline use with a user-friendly GUI.","language":"Python","stars":98,"topics":["ai-video-analysis","blip2","gui","image-captioning","image-captioning-ai","llm","object-detection","offline-processing","ollama","ollama-api"],"license":null,"category":"ai-agents","readme_excerpt":"AI-Powered Video Analyzer Offline, privacy-first AI video analysis. Runs entirely on your local machine — no cloud, no data upload, no telemetry. --- What it does Analyzes a video file through a local AI pipeline and produces structured outputs: Stage Technology Default --- --- --- Object detection VisionServeX D-FINE Required — primary backend Frame sampling Built-in (adaptive/scene-change) Enabled Scene captioning BLIP (Salesforce) Optional — needs [full] Speech transcription Whisper (OpenAI) Optional — needs [full] Audio events PANNs CNN14 Optional — needs [full] LLM summarization Ollama (any local model) Optional — needs Ollama All processing is local. Nothing is sent to any server. --- Quick start --- Installation Detection only (recommended starting point) Installs: frame sampler, VisionServeX D-FINE backend, CLI. No YOLO, no cloud dependencies. Full pipeline (all optional stages) Adds: Whisper, BLIP, PANNs, PyTorch, librosa, Ollama client. Development --- External tools ffmpeg — required for audio stages Ollama — optional, for LLM summarization --- First run --- Doctor command Checks: Python version, OpenCV, VisionServeX, D-FINE model registry, ffmpeg, PyTorch/GPU, Whisper, BLIP, PANNs, Ollama, moviepy, Tesseract. Required dependencies exit with ✗ . Optional dependencies show ✓ with an install hint when missing. Exits 0 if all required dependencies are present. --- Detection presets VisionServeX D-FINE models (COCO-80 classes, benchmarked on RTX 5080): Preset Model ms/","default_branch":null,"files":null,"tree":[],"storefront":"/r/arashsajjadi","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/arashsajjadi/ai-powered-video-analyzer/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}