An offline AI-powered video analysis tool with object detection (YOLO), image captioning (BLIP), speech transcription (Whisper), audio event detection (PANNs), and AI-generated summaries (LLMs via Ollama). It ensures privacy and offline use with a user-friendly GUI.
☆103Jun 14, 2026Updated 3 months ago
Alternatives and similar repositories for ai-powered-video-analyzer
Users that are interested in ai-powered-video-analyzer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A ComfyUI extension for generating captions of images.☆29May 12, 2025Updated last year
- Automatically generate a lip-synced avatar based off of a transcript and audio☆15Feb 17, 2023Updated 3 years ago
- Find and Use Cheats via the PythonGDB API☆11Aug 1, 2016Updated 10 years ago
- A lightweight MCP Server for GDB automation.☆15Aug 24, 2025Updated last year
- C++ library for audio and music analysis, description and synthesis, including Python bindings☆14Sep 4, 2026Updated 2 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆10May 28, 2020Updated 6 years ago
- Chrome extension that inserts a custom stylesheet and script into every web page.☆11Jul 16, 2019Updated 7 years ago
- ☆38Oct 6, 2025Updated 11 months ago
- A "loopback on steroids" type of extension for Stable Diffusion Web UI.☆31Oct 10, 2025Updated 11 months ago
- K3U Installer is a desktop GUI tool designed to simplify and automate the installation of ComfyUI with flexible setups☆14Apr 8, 2025Updated last year
- Epicycle .NET (C#) Photogrammetry Library☆12Nov 3, 2015Updated 10 years ago
- Upload a video and provide a prompt to generate a narration.☆13Mar 5, 2025Updated last year
- ☆16May 14, 2021Updated 5 years ago
- A simple luminance based image formation algorithm. Not optimal or particularly amazing, but a step in the right direction.☆13Jul 14, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A thumbnail gallery / set management tool for Automatic1111 Webui☆24Jul 31, 2024Updated 2 years ago
- Generating Video Caption Using LSTM☆12May 29, 2023Updated 3 years ago
- Public domain example files for every mimetype (in progress).☆14Feb 13, 2026Updated 7 months ago
- Implementation of Mask R-CNN architecture, one of the object recognition architectures, on a custom dataset.☆10Nov 1, 2022Updated 3 years ago
- A program to easily create and edit color matrices. They can then be used with NegativeScreen☆17Nov 20, 2016Updated 9 years ago
- Python Scripting is an integration of the open-source Python for .NET project. This package provides the means to import the Python Runti…☆16Oct 8, 2024Updated last year
- We implement RCNN algorithm for object detection from an Images.☆17Jul 6, 2020Updated 6 years ago
- One MCP server that gives Claude, ChatGPT, Cursor and any other AI app transcripts, chapters, metadata and frames from YouTube and 10 mor…☆22Updated this week
- This extension provides inference-time optimization techniques to enhance diffusion-based image generation quality through random search …☆23Feb 27, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ComfyUI nodes for RoyalCities Foundation-1 — structured text-to-sample music generation with full BPM, key, timbre, and FX control☆19Mar 17, 2026Updated 6 months ago
- Text to speech examples in Unity.☆15Feb 28, 2023Updated 3 years ago
- Voxella 🌍 - AI Video Translation and Dubbing App: Seamlessly translate and dub videos into multiple languages with Voxella. This powerfu…☆16Jun 4, 2023Updated 3 years ago
- Scaled Uniform Noise for Ancestral & Stochastic samplers and Noisy latent image☆17Mar 30, 2025Updated last year
- Integrating all DeepSeek open-source projects into ComfyUI, looking forward to DeepSeek’s OpenSourceWeek next week.☆17Feb 21, 2025Updated last year
- Constructed a dashboard with FastAPI that extracts data from the yfinance API to a SQLAlchemy database.☆21Mar 16, 2025Updated last year
- negamax AI algorithm for turn-based games☆13Oct 6, 2019Updated 6 years ago
- Kotlin extension for VSCode☆35Jun 9, 2018Updated 8 years ago
- Compare multiple Stable Diffusion models quickly with a contact sheet of multiple images from multiple models with the same prompt☆19Mar 11, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Running Deepseek R1 Local on Ollama☆28Jan 28, 2025Updated last year
- An advanced AI-powered tool that automatically translates and dubs YouTube videos into different languages while dynamically adjusting vi…☆18Nov 9, 2024Updated last year
- Collection of various text datasets to assist ML researchers in training or fine-tuning their models☆21Apr 1, 2023Updated 3 years ago
- 用open-cv检测物体的大小,是实时的☆17Jan 7, 2022Updated 4 years ago
- The easiest way to train a wan2.2 Lora.☆57Jul 28, 2026Updated last month
- Dummy project to test your Open3D build☆10May 6, 2021Updated 5 years ago
- Local GLM-4 Prompt Enhancer and Inference for ComfyUI☆31Jul 20, 2025Updated last year